Ensure availability, performance, scalability, and reliability of production systems. Monitor systems with observability tools, define SLIs/SLOs, lead incident management and RCAs, automate operations via IaC, and support CI/CD pipeline stability and improvements.
- Job Summary
- The Site Reliability Engineer (SRE) is responsible for ensuring the availability, performance, scalability, and reliability of enterprise platforms and applications. The role focuses on monitoring, automation, incident management, and continuous improvement, working closely with engineering and DevOps teams to build resilient and highly available systems.
- 2. Key Responsibilities
- Ensure high availability and reliability of production systems and services
Monitor system health using observability tools (logs, metrics, traces)
Define and track SLIs, SLOs, and SLAs to measure system performance
Lead/support incident management, root cause analysis (RCA), and post-incident reviews
Automate operational tasks and implement Infrastructure as Code (IaC) practices
Support and improve CI/CD pipelines for stable and efficient releases
3. Skills & Competencies - Technical Skills
- Cloud Platforms: Azure
Monitoring & Observability: Prometheus, Grafana, Splunk, ELK, Datadog
Containers & Orchestration: Docker, Kubernetes
CI/CD Tools: Jenkins, GitHub Actions, GitLab CI, Azure DevOps
Infrastructure as Code: Terraform, Ansible, CloudFormation
- 2. Key Responsibilities
- Ensure high availability and reliability of production systems and services
Monitor system health using observability tools (logs, metrics, traces)
Define and track SLIs, SLOs, and SLAs to measure system performance
Lead/support incident management, root cause analysis (RCA), and post-incident reviews
Automate operational tasks and implement Infrastructure as Code (IaC) practices
Support and improve CI/CD pipelines for stable and efficient releases
3. Skills & Competencies
- Technical Skills
- Cloud Platforms: AWS / Azure / GCP
Monitoring & Observability: Prometheus, Grafana, Splunk, ELK, Datadog
Containers & Orchestration: Docker, Kubernetes
CI/CD Tools: Jenkins, GitHub Actions, GitLab CI, Azure DevOps
Infrastructure as Code: Terraform, Ansible, CloudFormation
Part of the $4.8 billion RPG Group, we’re a community of 10,000+ innovators across 30+ global locations, including Milpitas, Seattle, Princeton, Cape Town, London, Zurich, Singapore, and Mexico City. Explore Life at Zensar and join us to Grow. Own. Achieve. Learn. to be the best version of yourself.
We believe the best work happens when individuality is celebrated, growth is encouraged, and well-being is prioritized. We are an equal employment opportunity (EEO) and affirmative action employer, committed to creating an inclusive workplace. All qualified applicants will be considered without regard to race, creed, color, ancestry, religion, sex, national origin, citizenship, age, sexual orientation, gender identity, disability, marital status, family medical leave status, or protected veteran status.
Zensar Technologies Pune, Mahārāshtra, IND Office
Zensar Knowledge Park, Kharadi, Plot # 4, MIDC, Pune, Maharashtra, India, 411014
Zensar Technologies Pune, Maharashtra, IND Office
Pune, India
Similar Jobs
Information Technology
Design, operate, and automate AWS-based production and development infrastructure for eCommerce/enterprise platforms. Implement CI/CD, infrastructure-as-code, observability, security, disaster recovery, and large-scale automation. Support microservices, databases, caching, virtualization, and application servers while troubleshooting, performance tuning, and onboarding new tools.
Top Skills:
AkamaiAndroidApacheAppdynamicsAptitude/DpkgAwkAWSBashCassandraCdnChefDatadogDynatraceEc2Elastic CloudElkGeodnsGlobal Traffic ManagementGraphiteGroovyHadoopHaproxyHbaseHelmHypervisorIptablesJavaJavaScriptJbossJenkinsJettyJSONKeycloakKubernetesLdapLinuxMemcachedMicroservicesMongoDBMySQLNagiosNessusNetappNew RelicNfsNginxNmapNtpObjective-COktaOpen DirectoryOraclePerlPHPPuppetPythonRackspace CloudRedisRestRubyService MeshSoftlayerSplunkSsl/TlsTerraformTomcatVarnishVdiVirtualizationWeblogicXMLYum/Rpm
Information Technology
Design, implement, and maintain scalable, secure GCP infrastructure and Kubernetes deployments. Build and optimize CI/CD pipelines (Jenkins), automate IaC (Terraform/Deployment Manager), monitor systems (Prometheus, Grafana, Cloud Monitoring, ELK), troubleshoot production issues, apply SRE practices (SLIs/SLOs, incident management), and harden Linux-based environments.
Top Skills:
ArgocdBashCloud MonitoringCloud StorageCompute EngineDeployment ManagerDockerElasticsearchElkGCPGitGkeGoGrafanaIamJbossJenkinsKibanaKubernetesLinuxLogstashPrometheusPythonSpinnakerStackdriverTerraformVaultVpcWildfly
Security • Software
Design, build, and automate CI/CD and deployment toolsets for cloud-hosted SaaS microservices. Deploy and configure containerized applications to Tomcat and Kubernetes, manage MongoDB test databases, troubleshoot build and environment issues, own GitHub/Artifactory/Jenkins release activities, optimize cloud infrastructure costs, and participate in a 24/7 on-call rotation while collaborating with development, architecture, and product teams.
Top Skills:
ArtifactoryAws CloudwatchAws Ec2Aws EcsAws EksAws FargateAws IamAws S3Aws SnsAws SqsBitbucketBlackduckCi/CdContainersDockerGitGradleGroovyJavaJenkinsKubernetesLinuxMavenMongoDBNpmPythonSonarTerraformTomcat
What you need to know about the Pune Tech Scene
Once a far-out concept, AI is now a tangible force reshaping industries and economies worldwide. While its adoption will automate some roles, AI has created more jobs than it has displaced, with an expected 97 million new roles to be created in the coming years. This is especially true in cities like Pune, which is emerging as a hub for companies eager to leverage this technology to develop solutions that simplify and improve lives in sectors such as education, healthcare, finance, e-commerce and more.

