GoGuardian Logo

GoGuardian

Senior Site Reliability Engineer

Posted 2 Days Ago
Be an Early Applicant
Remote
Hiring Remotely in India
Senior level
Remote
Hiring Remotely in India
Senior level
Designs, scales, and maintains highly available cloud infrastructure for production SaaS applications. Responsibilities include improving observability and alerting, leading incident response and postmortems, modernizing CI/CD pipelines, managing infrastructure as code, supporting data layers, partnering with engineering teams, and enforcing security and compliance standards across cloud environments.
The summary above was generated by AI
What We Do
 
At GoGuardian, we’re helping build a future where all learners are ready and inspired to solve the world’s greatest challenges. Our award-winning system of learning solutions is purpose-built for K-12 and trusted by school leaders to promote effective teaching and equitable engagement while helping empower educators to keep students safe. 
 
What It’s Like to Work at GoGuardian

We are an outcomes-focused learning company with a steadfast focus on improving learning environments, one classroom at a time. Working with us means joining a remote team of diverse, committed, mission-driven employees who are inspired by our vision, dedicated to our customers, and ready to roll up their sleeves. Guardians put their heads together to solve problems, learn together from experiments that fail, and stand together by their work with full accountability. We balance our diligence with an inclusive culture that invites everyone to bring their whole self to work. Join us and learn why “I love the people here” is one of the most frequent comments we hear from Guardians.

What We Do

At GoGuardian, we’re helping build a future where all learners are ready and inspired to solve the world’s greatest challenges. Our award-winning system of learning solutions is purpose-built for K-12 and trusted by school leaders to promote effective teaching and equitable engagement while helping empower educators to keep students safe. 


The Role

We’re looking for a Senior Site Reliability Engineer (SRE) to help design, scale, and maintain the infrastructure that powers our core products and services. In this role, you’ll collaborate with engineering teams to drive operational excellence, optimize system performance, and ensure high availability across production environments. This position sits on Tech Foundation, a team that manages core cloud infrastructure, shared data services, and developer tooling to empower our product teams to deliver software efficiently and securely. The ideal candidate brings a strong background in cloud infrastructure, automation, and modern reliability practices, with a passion for solving complex operational challenges in a collaborative environment.


What You'll Do

  • Architect and maintain scalable, secure cloud infrastructure to ensure high availability for core products.
  • Enhance observability and monitoring frameworks to deliver highly accurate alerts, minimizing noise and improving incident detection.
  • Participate in on-call rotations and lead incident response, ensuring comprehensive post-mortems and RCAs are completed to drive systemic improvements.
  • Optimize and modernize deployment pipelines and automation workflows to maximize engineering velocity and operational safety.
  • Partner with product development teams to provide infrastructure support, review architectural changes, and promote reliability best practices.
  • Implement and uphold robust security standards and compliance controls across all managed cloud infrastructure.


Who You Are

  • 5+ years of professional experience in Site Reliability Engineering, Infrastructure, or DevOps roles supporting production SaaS applications.
  • Strong proficiency with AWS core services (including EC2, VPC, S3) along with experience in Serverless frameworks and managed Kubernetes environments like EKS.
  • Extensive experience writing and managing Infrastructure as Code (IaC) using Terraform.
  • Familiarity with configuring, troubleshooting, and maintaining data layers such as MongoDB, Redshift, and OpenSearch. Experience with GCP environments or technologies like Firestore is a plus.
  • Experience managing or modernizing CI/CD pipelines and deployment workflows utilizing systems like Jenkins, AWS CodeBuild/CodePipeline, or GitHub Actions.
  • Deep understanding of Linux operating system fundamentals and Unix shell scripting.
  • Ability to read and debug code written in JavaScript/TypeScript, Python, or Go to effectively troubleshoot underlying service errors.
  • Strong communication and collaboration skills, with a track record of driving technical decisions and establishing team-wide operational standards.
  • Eager to take initiative in a fast-paced, ever-changing, dynamic environment.
  • Fueled by the opportunity to truly impact the education landscape.
  • Something else? Tell us! We want to learn more about you…


What It’s Like to Work at GoGuardian

We are an outcomes-focused learning company with a steadfast focus on improving learning environments, one classroom at a time. Working with us means joining a remote team of diverse, committed, mission-driven employees who are inspired by our vision, dedicated to our customers, and ready to roll up their sleeves. Guardians put their heads together to solve problems, learn together from experiments that fail, and stand together by their work with full accountability. We balance our diligence with an inclusive culture that invites everyone to bring their whole self to work. Join us and learn why “I love the people here” is one of the most frequent comments we hear from Guardians.

 What We Offer 

  • Competitive pay, health insurance, accident insurance, life insurance, and a retirement savings plan.
  • Paid annual leave, paid holidays, paid parental leave, paid leave for life events, and a paid year-end holiday break.
  • A robust catalog of benefits that support your professional growth and personal wellbeing, including work from home funds, wellness checks, and more…

Plus the intangible:

  • A varied and challenging role in an innovative, global company.
  • Supportive, driven colleagues who have your back and share your passion.

GoGuardian is an equal opportunity employer and makes employment decisions on the basis of merit and business needs. GoGuardian does not discriminate against employees, applicants, interns, or volunteers on the basis of race, religion, color, national origin, ancestry, physical disability, mental disability, medical condition, pregnancy, marital status, sex, age, sexual orientation, military and veteran status, registered domestic partner status, genetic information, gender, gender identity, gender expression, or any other characteristic protected by applicable law.


Please share this with your friends or co-workers who may be interested in working at GoGuardian! We have multiple openings and are always looking for talented people. 

 
GoGuardian is an equal opportunity employer and makes employment decisions on the basis of merit and business needs. GoGuardian does not discriminate against employees, applicants, interns or volunteers on the basis of race, religion, color, national origin, ancestry, physical disability, mental disability, medical condition, pregnancy, marital status, sex, age, sexual orientation, military and veteran status, registered domestic partner status, genetic information, gender, gender identity, gender expression, or any other characteristic protected by applicable law.
 
GoGuardian's Job Applicant Privacy Policy is located here. 
 
#BI-Remote

Similar Jobs

18 Days Ago
Remote or Hybrid
Senior level
Senior level
Digital Media • eCommerce • Gaming • Mobile • News + Entertainment
Lead reliability, scalability, observability, automation, infrastructure, disaster recovery, and security initiatives for Crunchyroll’s cloud-native data platforms. Establish SRE practices including SLIs, SLOs, error budgets, incident management, and postmortems. Operate Kubernetes and GCP environments, implement Infrastructure as Code, optimize capacity and performance, and drive vulnerability remediation, penetration-testing support, and cloud platform security.
Top Skills: Ci/CdDatadogGCPGoGrafanaIdentity And Access ManagementInfrastructure As CodeJavaKubernetesLinuxOpentelemetryOwasp Top 10PrometheusPythonShellTerraform
4 Days Ago
In-Office or Remote
Senior level
Senior level
Software
Owns Kubernetes-based development, CI, pre-production, and customer-facing production environments. Responsibilities include SRE operations, incident response, on-call support, Helm and CI/CD lifecycle management, infrastructure automation, observability, security, disaster recovery, stateful platform services, and multi-region deployments. The role also provides technical consultation, creates operational documentation, validates upgrades, and mentors engineers.
Top Skills: AnsibleApi GatewaysArgo CdAWSCluster ApiDockerFluxGithub ActionsGitopsGoGrafanaHelmKafkaKeycloakKindKubernetesKyvernoMetal3OpaOpenstackOpentelemetryPostgresPrometheusPytestPythonTemporalTerraform
4 Days Ago
In-Office or Remote
Senior level
Senior level
Internet of Things • Mobile • Retail
Own platform reliability, DevOps automation, CI/CD, Kubernetes deployments, observability, logging, cloud infrastructure governance, and incident response. Lead Kafka and streaming-platform operations, capacity and disaster-recovery planning, access controls, cost governance, and reliability improvements. Provide advanced production troubleshooting for microservices and mentor Tier 1 and Tier 2 support teams.
Top Skills: AlertmanagerApache AirflowApache FlinkAws MskAzureAzure Container Registry (Acr)Azure Event HubsAzure Kubernetes Service (Aks)Azure MonitorCi/CdConfluent CloudConfluent KafkaFluent BitGithub ActionsGrafanaHelmJavaJfrogKubernetesOpensearchPostgresPrometheusPythonReactSpring BootThanos

What you need to know about the Pune Tech Scene

Once a far-out concept, AI is now a tangible force reshaping industries and economies worldwide. While its adoption will automate some roles, AI has created more jobs than it has displaced, with an expected 97 million new roles to be created in the coming years. This is especially true in cities like Pune, which is emerging as a hub for companies eager to leverage this technology to develop solutions that simplify and improve lives in sectors such as education, healthcare, finance, e-commerce and more.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account