- Goguardian
- India, MP
- <72 Hours
Site Reliability Engineer II.
Before you go
Before you leave us, sign up for our email alerts
We don't do job spam, just the best digital jobs delivered straight to your inbox.
Site Reliability Engineer II: our view in 3 lines...
- The Role:This role is for a Site Reliability Engineer II supporting cloud infrastructure and production services for a K-12 learning company.
- The Person:The person will build and support cloud infrastructure, maintain monitoring and alerting, join on-call rotations, handle incident response and post-mortems, improve deployment automation, and apply security and compliance controls.
- Requirements:The ideal candidate has 2+ years of experience in Site Reliability Engineering, Infrastructure, or DevOps roles, with working knowledge of AWS, Terraform, Kubernetes, CI/CD tools, Linux, and Unix shell scripting.
About the role
We are an outcomes-focused learning company with a steadfast focus on improving learning environments, one classroom at a time. Working with us means joining a remote team of diverse, committed, mission-driven employees who are inspired by our vision, dedicated to our customers, and ready to roll up their sleeves. Guardians put their heads together to solve problems, learn together from experiments that fail, and stand together by their work with full accountability. We balance our diligence with an inclusive culture that invites everyone to bring their whole self to work. Join us and learn why “I love the people here” is one of the most frequent comments we hear from Guardians.
The Role
We’re looking for a Site Reliability Engineer (SRE) II to help build, maintain, and scale the infrastructure that powers our core products and services. In this role, you’ll work alongside engineering teams to support operational excellence, optimize system performance, and ensure high availability across production environments. This position sits on Tech Foundation, a team that manages core cloud infrastructure, shared data services, and developer tooling to empower our product teams to deliver software efficiently and securely. The ideal candidate has practical experience with cloud infrastructure, automation, and core reliability practices, with a strong desire to solve complex operational challenges in a collaborative environment.
_________________________________________________________________________________________
What You'll Do
- Build, maintain, and support scalable cloud infrastructure to ensure high availability for core products.
- Maintain and enhance observability and monitoring frameworks to deliver accurate alerts and improve incident detection.
- Participate in on-call rotations and support incident response, helping conduct post-mortems and trace RCAs to mitigate future issues.
- Maintain and improve deployment pipelines and automation scripts to support operational safety and engineering velocity.
- Collaborate with product development teams to provide day-to-day infrastructure support and assist with operational best practices.
- Apply established security standards and compliance controls across all managed cloud infrastructure.
Who You Are
- 2+ years of professional experience in Site Reliability Engineering, Infrastructure, or DevOps roles supporting production SaaS applications.
- Working knowledge of AWS core services (including EC2, VPC, S3) along with exposure to Serverless frameworks or managed Kubernetes environments like EKS.
- Hands-on experience writing and maintaining Infrastructure as Code (IaC) using Terraform.
- Familiarity with supporting, troubleshooting, or monitoring data layers such as MongoDB, Redshift, or OpenSearch.
- Experience working with CI/CD tools and deployment workflows using systems like Jenkins, AWS CodeBuild/CodePipeline, or GitHub Actions.
- Solid understanding of Linux operating system fundamentals and Unix shell scripting.
- Ability to read and debug code written in JavaScript/TypeScript, Python, or Go to assist in troubleshooting application-level errors.
- Strong communication and interpersonal skills, with a collaborative approach to solving technical problems across teams.
- Eager to take initiative in a fast-paced, ever-changing, dynamic environment.
- Fueled by the opportunity to truly impact the education landscape.
- Something else? Tell us! We want to learn more about you…
Please share this with your friends or co-workers who may be interested in working at GoGuardian! We have multiple openings and are always looking for talented people.
