Site Reliability Engineer II

Apply Now โ†—
๐ŸŒ Remote๐Ÿ“ Costa Rica

About this role

Do you like collaborating across teams to solve complex problems?

Do you enjoy solving large scale distributed content delivery challenges?

Join our highly skilled Site Reliability Team

The Platform & Reliability Engineering team is responsible for defining, measuring, & optimizing the key performance indicators of delivery customers. Your expertise in software engineering and systems administration will be instrumental in building robust and resilient infrastructure.

Partner with the best

In this role, you'll play a pivotal role in shaping the future of our products. You'll collaborate closely with product teams to ensure the reliability, scalability, and performance of our systems. You'll define key performance indicators (KPIs). Advance the state of monitoring, alerting and operational responses, and investigate complex performance issues.

As a Site Reliability Engineer II, you will be responsible for:

  • Working on Internet technologies to improve the performance, availability, and scalability of large distributed content delivery systems
  • Engaging in collaborative efforts with cross-functional teams to define and establish measurable Service Level Indicators and Service Level Objectives
  • Monitoring platform availability and performance, debug issues by leveraging data analysis skills and implement corrective actions to avoid recurrence
  • Developing and implement automation solutions to improve operational efficiency and reduce toil.
  • Improving CI/CD pipelines and safe deployment practices for platform services.
  • Participating in design reviews and providing technical guidance to ensure designs meet requirements for scalability, performance, and robustness

Do what you love

To be successful in this role you will:

  • Have 2 years of relevant experience and a Bachelor's degree in Computer Science or its equivalent
  • Have hands-on experience with compute platforms such as Kubernetes, Containerization, and Docker
  • Have experience with monitoring and alerting systems (e.g., Prometheus, Grafana, ADBMS, Datadog), including metric collection, alerting, dashboarding, and troubleshooting
  • Show fluency working in a UNIX/Linux computing environment
  • Have familiarity with infrastructure-as-code tools such as Terraform
  • Have proficiency with a configuration management tool such as Ansible, Salt Stack, Chef, Puppet, or similar

About us

At Akamai, we make life better for billions of people, trillions of times a day.
Whether you're streaming live events, scrolling social media, watching your favorite series, or managing your savings, we're the engine behind the scenes. We provide the world's most distributed platform from Cloud to Edge to help the giants of the digital world work faster and stay more secure, making the internet a better experience for everyone.

Our focus is simple:
Cloud and Edge: Running apps closer to users for instant performance.
Security: Neutralizing threats before they ever reach your data.
Content Delivery: Scaling the world's biggest moments without a glitch.
AI: Enabling our customers to build, secure, and scale AI apps on the world's most distributed cloud platform.

At Akamai, we don't just support the internet; we power and protect it, because behind every great digital experience is a massive hidden challenge. And we're the ones who solve it. When millions of people hit play or pay, Akamai ensures it just works.

Benefits at Akamai: We support your health, well-being, finances, and life beyond work. See our benefits.

FlexBase adapts to your job's needs

Akamai's FlexBase program is yet another way we show our commitment to providing employees with an exceptional workplace experience. It's not about telling employees where to work; it's about supporting employees to do their best work.

We trust our incredible employees to work in ways that suit them best: at home, in an office, or a combination of both.

Connect with us on social and see what life at Akamai is like!

Frequently Asked Questions

Is the salary disclosed for the Site Reliability Engineer II position at Akamai?
The salary for this Site Reliability Engineer II role at Akamai is not publicly listed. Click "Apply Now" to learn more about the compensation package on their official careers page.
Is the Site Reliability Engineer II job at Akamai remote?
Yes, this Site Reliability Engineer II position at Akamai is remote, with team members based in Costa Rica. You can work from home or anywhere in the supported regions.
How do I apply for the Site Reliability Engineer II position at Akamai?
Click the "Apply Now" button on this page. You will be redirected to Akamai's official application portal hosted on oraclecloud where you can submit your application directly.
When was the Site Reliability Engineer II job at Akamai posted?
This Site Reliability Engineer II position at Akamai was posted on Jul 27, 2026. Apply as soon as possible โ€” early applications are often reviewed first.
Site Reliability Engineer II
Akamai
Apply for this role โ†—

You'll be redirected to Akamai's official application page on oraclecloud.