Akamai Technologies Logo

Akamai Technologies

Site Reliability Engineer II

Posted Yesterday
Be an Early Applicant
In-Office or Remote
Hiring Remotely in India
Junior
In-Office or Remote
Hiring Remotely in India
Junior
Designs and builds automated operational workflows for Akamai’s global backbone and network infrastructure. Responsibilities include reducing manual toil, improving observability and telemetry, developing maintainable software with CI/CD and test-driven practices, and deploying cloud-native third-party systems using Infrastructure as Code. The role requires programming, scalable full-stack development, Kubernetes, microservices, DevOps, database, monitoring, dashboarding, and alerting experience.
The summary above was generated by AI

Do you have a passion for network automation?

Do you want to help scale a global platform?

Be part of Akamai's Core IP Networking team!

Our team supports Akamai's global backbone and networking technologies by solving operational challenges through code. This position focuses on reducing toil through code, improving our ability to build and scale. As SRE, you will be a crucial part of our ability to deliver global infrastructure connectivity for the next generation of Akamai's services.

Partner with the best

Our SRE team sits directly next to Architecture, Engineering and Ops. We are tasked with identifying the biggest areas for improvement and are responsible for designing and building solutions. With a focus on software, you will be directly responsible for continuously improving Akamai's ability to monitor and operate our backbone.

As a Site Reliability Engineer II, you will be responsible for:

  • Designing and building systems that support automated operational workflows and provide deep insights into our global network data
  • Identifying and eliminate bottlenecks in operational workflows caused by process inefficiencies, manual toil, and technical debt
  • Generating telemetry data and deep observability for network infrastructure
  • Architecting our code to be self-sustainable, robust and easily maintainable, focusing on CI/CD, test-driven development, and iterative design
  • Supporting and deploying cloud-native, third party software utilizing an infrastructure as code paradigm

Do what you love

To be successful in this role you will:

  • Have 2 and more years of relevant experience with a Bachelor's degree in Computer Science, Engineering, or a related field,
  • Have programming experience with Python/JavaScript, ability to design and develop scalable full-stack systems using appropriate frameworks and SDLC best practices.
  • Have experience with containers, Kubernetes, service orchestration, microservices, Infrastructure as Code (IaC), automation, and modern DevOps practices.
  • Have experience with PostgreSQL, Apache Cassandra, network telemetry, monitoring/metrics, dashboarding, alerting, and developing automation to improve operational efficiency.

About us

At Akamai, we make life better for billions of people, trillions of times a day.
Whether you're streaming live events, scrolling social media, watching your favorite series, or managing your savings, we're the engine behind the scenes. We provide the world's most distributed platform from Cloud to Edge to help the giants of the digital world work faster and stay more secure, making the internet a better experience for everyone.
Our focus is simple:
Cloud and Edge: Running apps closer to users for instant performance.
Security: Neutralizing threats before they ever reach your data.
Content Delivery: Scaling the world's biggest moments without a glitch.
AI: Enabling our customers to build, secure, and scale AI apps on the world's most distributed cloud platform.
At Akamai, we don't just support the internet; we power and protect it, because behind every great digital experience is a massive hidden challenge. And we're the ones who solve it. When millions of people hit play or pay, Akamai ensures it just works.

Benefits at Akamai: We support your health, well-being, finances, and life beyond work. See our benefits.

FlexBase adapts to your job's needs

Akamai's FlexBase program is yet another way we show our commitment to providing employees with an exceptional workplace experience. It's not about telling employees where to work; it's about supporting employees to do their best work.
We trust our incredible employees to work in ways that suit them best: at home, in an office, or a combination of both.

Connect with us on social and see what life at Akamai is like!

Similar Jobs

4 Hours Ago
Remote
India
Senior level
Senior level
Software
The Site Reliability Engineer will design scalable infrastructure, automated deployment pipelines, monitoring, and alerting systems. Responsibilities include troubleshooting production incidents, participating in on-call rotations, maintaining systems, improving security and compliance, optimizing reliability and efficiency, and mentoring junior engineers. The role requires collaboration with development and cross-functional teams, technical communication, infrastructure automation, and expertise in distributed systems, cloud platforms, and networking.
Top Skills: AnsibleAWSAzureChefDockerGCPGoGrafanaJavaKubernetesNagiosPrometheusPuppetPythonRubyTerraform
6 Hours Ago
In-Office or Remote
India
Junior
Junior
Cloud • Security • Software • Cybersecurity
Build and improve reliable, scalable distributed content delivery systems. Define SLIs and SLOs, enhance monitoring and alerting, analyze performance data, resolve complex incidents, automate operational tasks, and participate in architecture reviews. The role requires scripting, Oracle SQL analysis, Unix/Linux expertise, and experience with observability tools including Prometheus, Grafana, ADBMS, and Datadog. Collaboration with product and cross-functional teams is central to ensuring high availability, performance, and resilience.
Top Skills: AdbmsBashCloud ComputingDatadogDevOpsGrafanaJavaScriptOracle SqlPrometheusPythonUnix/Linux
2 Days Ago
Remote
India
Junior
Junior
Edtech • Software
Build and maintain scalable AWS cloud infrastructure, observability and monitoring systems, deployment pipelines, and automation. Participate in on-call rotations, incident response, post-mortems, and root-cause analysis. Support product engineering teams with infrastructure troubleshooting and operational best practices while applying security and compliance controls. The role requires experience with Terraform, CI/CD, Linux, shell scripting, cloud services, managed Kubernetes, databases, and debugging application code.
Top Skills: Amazon RedshiftAWSAws CodebuildAws CodepipelineEc2EksGithub ActionsGoJavaScriptJenkinsLinuxMongoDBOpensearchPythonS3Serverless FrameworksTerraformTypescriptUnix ShellVpc

What you need to know about the Pune Tech Scene

Once a far-out concept, AI is now a tangible force reshaping industries and economies worldwide. While its adoption will automate some roles, AI has created more jobs than it has displaced, with an expected 97 million new roles to be created in the coming years. This is especially true in cities like Pune, which is emerging as a hub for companies eager to leverage this technology to develop solutions that simplify and improve lives in sectors such as education, healthcare, finance, e-commerce and more.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account