Verified today
Senior Site Reliability Engineer
About the role
Gurugram Products & Technology team at MongoDB builds the operational foundation for a new AI application platform. The Senior Site Reliability Engineer designs, scales, and maintains multi‑tenant Kubernetes infrastructure, ensures reliability, and leads on‑call incident response. Based in Gurugram with a hybrid work model, the role mentors early‑career SREs and drives automation across cloud environments.
What you’ll do
- Operate and improve multi‑tenant Kubernetes fleet
- Build reliable, self‑healing services and infrastructure
- Define metrics for incident detection and service health
- Participate in 24/7 on‑call rotation
- Mentor early‑career SREs
- Collaborate with product and engineering teams
- Implement automation for deployment and scaling
- Contribute to operational best practices
What you’ll bring
- 6+ years building and operating distributed systems
- Proficiency in Python or Go
- Production Kubernetes experience
- Cloud expertise (AWS, GCP, or Azure)
- Strong Linux and networking knowledge
- Experience with observability and alerting
- Ability to mentor junior engineers
- Bias for automation and operational simplicity