Skip to main content

Senior Site Reliability Engineer, Fleet Management

MongoDB
Remote - USUpdated 10d ago
Base salary
$127k–$249k
Published base salary range
Location
Remote - US
Remote eligibility
Employment
Full-time
Senior
Role family
Infrastructure
Developer tools
Apply on mongodb.com
Job actionsApply now
Job actionsApply now

About the job

About the Role

MongoDB's Platform Engineering department within SRE is responsible for critical infrastructure and operational functions supporting the broader engineering organization, including multi-cloud Kubernetes infrastructure, networking, load balancing, and observability. The Fleet Management team provides the core runtime environment for developers, managing the end-to-end lifecycle of the Kubernetes fleet and critical components ensuring cluster reliability and security (e.g., CoreDNS, cert-manager, Gatekeeper). The team is spearheading a migration from Terraform-based Infrastructure as Code to an Operator-driven lifecycle management model.

Responsibilities

  • Contribute to developing and maintaining a scalable and secure runtime environment on top of Kubernetes that supports product needs across MongoDB.
  • Provide internal support for the Kubernetes ecosystem, partnering with engineering teams to solve domain-specific problems.
  • Participate in a 24/7 on-call rotation to resolve critical issues.
  • Prioritize blameless post-mortems and dedicate engineering time to systemic fixes.

Qualifications

  • 6+ years of experience in software development and operating distributed systems.
  • Proficient in Go, Python, or similar language, with strong commitment to code quality and testing (unit, integration, E2E).
  • Deep experience using and extending containerization technologies, preferably Kubernetes.
  • Solid understanding of Linux internals and networking concepts (filesystems, TCP/IP, DNS, TLS).
  • Customer-focused mindset, treating internal developers as primary users.
  • Strong operational ownership, debugging complex production issues.
  • Prefer automation over manual processes.

Strong Candidates May Also Have

  • Experience designing secure, multi-tenant runtime environments from first principles.
  • Proficiency with Kubernetes ecosystem tools: Helm, Kustomize, Gatekeeper, Kyverno, CRDs/Operators, CRI, CSI.
  • Expertise in cloud platforms: AWS, GCP, or Azure.
  • Proficiency with Terraform, Crossplane, AWS Controllers for Kubernetes (ACK).
  • Advanced Linux internals and networking concepts for containers (namespaces, cgroups).

Special Requirement

Must be a US Citizen.

Compensation & Benefits

Base salary range for this role in the U.S. is $127,000—$249,000 USD. Other benefits for eligible U.S.-based employees may include: equity, employee stock purchase program, flexible paid time off, 20 weeks fully-paid gender-neutral parental leave, fertility and adoption assistance, 401(k) plan, mental health counseling, transgender-inclusive health insurance coverage, and health benefits offerings.

Application Instructions

Apply via the provided link.

Skills & tags

What you can verify before applying

Compare the essentials before you leave: pay, remote scope, employment type, source, and the employer apply destination.