$ cat jobs/staff-site-reliability-engineer-mongodb-f660ee165452.json
Staff Site Reliability Engineer
We are seeking a Staff Site Reliability Engineer to join our growing Gurugram Products & Technology team to provide technical direction, shape architecture, and build key operational foundations of a new platform we are building to make it easier for customers to build AI applications using MongoDB. As a Staff Site Reliability Engineer on this new team, you will be responsible for providing technical leadership for the operational foundations that enable deployment at scale of AI applications. You will own the reliability architecture of the platform as it expands across regions and cloud providers, and set the technical direction for how the platform is operated, including capacity planning, multi-cloud expansion, incident response, and SLO discipline. The platform's SRE team owns the operational foundations: the Kubernetes fleet, networking, observability and alerting, and tenant isolation. MongoDB engineering teams pride themselves on building high-quality software and living MongoDB cultural values every day – we value intellectual curiosity and honesty, and building together in an environment that prioritizes collaboration over competition. We are looking to speak to candidates who are based in Bengaluru for our hybrid working model. Position Expectations Own the reliability architecture of the platform across regions and cloud providers Collaborate with the teams building the platform, providing internal support and guidance on operability, capacity, and best practices Set operational standards for the team: on-call quality, incident response, SLO discipline Mentor and technically develop the SRE team Participate in a 24/7 on-call rotation to resolve issues involving platform infrastructure Qualifications 10+ years of experience working on software and operating distributed systems, with deep Kubernetes expertise, including designing or evolving multi-cluster platforms. Proficiency in Python, Go, or a similar programming language Understand workload isolation
Similar remote roles
Senior Site Reliability Engineer
MongoDB · Worldwide · senior
Platform Database Engineer (MONGO DB)
valtech · LATAM · mid
Platform Database Engineer
valtech · LATAM · mid
Site Reliability Engineer (SRE/ DevOps) - Engineering Productivity
Arista Networks · Worldwide · mid
Senior/Staff/Lead SRE and Infra SWE
MongoDB · Worldwide · senior
Data Engineer
SumUp · EU · mid
Senior SRE/DevOps Engineer (Remote, Colombia/ Brazil) - Min.
exceptionly · US · senior
Senior Lead Database Reliability Engineer
draftkings inc · US · senior