01Responsibilities
Design, develop, ship, and motivate the creation of software and systems to increase product reliability and organizational efficiency.
Guide reliability practices through the entire software development lifecycle through activities like architecture reviews, code reviews, creating platforms and frameworks, capacity planning, and chaos testing.
Maintain service health through monitoring and incident response.
Improve service reliability through blameless post-incident reviews and using code to prevent or respond to problem recurrence.
Required Skills:
BS degree in Computer Science, related technical field, or equivalent practical experience.
Experience writing code in Java, Go, Shell, Perl, Python, or a similar language.
Ability to debug, optimize code, and automate routine tasks.
Strong interest in SRE topics like SLOs, resilience, scaling, performance, and more.
We get excited about engineers who:
have previous experience with public clouds and enabling technology (AWS & Terraform preferred or other similar technologies)
Familiar on one or more CI/CD Tools Maven, Gradle, Fastlane, Jenkins, Travis/CircleCI
Experience designing using asynchronous messaging patterns .