🏢
MongoDB
Site Reliability Engineer (Senior or Staff), Atlas
Job Description
ABOUT THE ROLE
MongoDB’s Atlas platform is the world’s most widely deployed, globally distributed, multi‑cloud data platform. As a Senior or Staff Site Reliability Engineer on the Atlas SRE team, you will be at the forefront of building and operating the infrastructure that powers millions of customer applications, from AI‑native startups to Fortune 100 enterprises. Your work will directly influence the reliability, performance, and scalability of Atlas, ensuring that every customer can run their most critical workloads with confidence. Whether you work from our New York City headquarters or from a location in the Eastern or Central time zone, you will collaborate with cross‑functional teams to design resilient systems, automate operations, and drive continuous improvement across a multi‑cloud environment.
WHAT YOU'LL DO
You will design and implement large‑scale, multi‑cloud solutions that underpin Atlas, leveraging AWS, Azure, and Google Cloud to deliver high availability and low latency for customers worldwide. You will own the full lifecycle of critical services, from architecture and deployment to monitoring, incident response, and post‑mortem analysis. Your day‑to‑day responsibilities will include automating repetitive operational tasks, developing tooling that reduces manual toil, and collaborating with software engineering teams to embed reliability best practices into the development pipeline. You will participate in a 24/7 on‑call rotation, diagnosing and resolving incidents that affect the Atlas fleet, and you will lead blameless post‑mortems to drive systemic improvements. Finally, you will mentor junior SREs, share knowledge across the organization, and contribute to the evolution of our reliability culture.
WHAT YOU'LL NEED
You bring at least five years of experience running critical systems at scale, with a proven track record of delivering reliable, high‑performance services in a production environment. You are comfortable operating a large‑scale Linux environment, understanding low‑level fundamentals, and you have deep expertise in at least one modern programming language such as Go, Python, or Ruby. Your background includes designing and managing infrastructure in a multi‑cloud setting, and you are familiar with core web and network protocols including HTTP, TLS, and DNS. You have a customer‑first mindset, prioritizing automation over manual processes, and you thrive in a fast‑paced, highly collaborative environment. U.S. citizenship is required for this role.
WHY REMOTE
MongoDB embraces a distributed, asynchronous culture that empowers engineers to work when they are most productive while maintaining a strong sense of collaboration. Remote employees enjoy flexible schedules that support work‑life balance, and the company provides the tools and processes necessary to stay connected across time zones. By removing geographic constraints, we attract top talent from around the world and foster a diverse, inclusive community that drives innovation in the data platform space.
BENEFITS
MongoDB offers a competitive benefits package that includes comprehensive health, dental, and vision coverage; a generous paid time off program; a 401(k) plan with company match; equity and employee stock purchase options; 20 weeks of fully paid, gender‑neutral parental leave; fertility and adoption assistance; and mental health counseling. Remote employees receive a stipend to support home‑office setup and an annual learning budget to pursue professional development. All benefits are designed to support the well‑being and growth of our team members, enabling them to thrive both personally and professionally.