Verified today
Senior Site Reliability Engineer
About the role
Platform & Reliability Engineering team at Akamai builds and operates the large‑scale distributed content delivery platform. The Senior Site Reliability Engineer designs, implements and maintains robust infrastructure, automation and monitoring to ensure reliability, scalability and performance. Works cross‑functionally with product and engineering to define SLOs, troubleshoot incidents and drive automation across the organization. Remote, India.
What you’ll do
- Design and implement scalable, reliable infrastructure for content delivery
- Develop automation to reduce manual operational tasks
- Define and monitor service level indicators and objectives
- Collaborate with product and engineering on reliability requirements
- Analyze incidents, perform root‑cause analysis and drive corrective actions
- Participate in design reviews ensuring performance and robustness
- Stay current with cloud, DevOps and SRE best practices
- Mentor junior engineers and share knowledge
What you’ll bring
- 5+ years relevant experience
- Bachelor's degree in Computer Science or related field
- Proficiency in Python, Bash and JavaScript scripting
- Experience with Prometheus, Grafana, Datadog monitoring
- Strong UNIX/Linux systems expertise
- Oracle SQL for data integrity checks
- Automation development and tool building
- Excellent communication and teamwork