Senior Site Reliability Engineer, Fleet Management

MongoDB

Confirmed live yesterday High trust
Remote

Quick summary

Work type
Remote
Location
Austin, TXBoston, MALos Angeles, CANew York, NYRaleigh, NCSan Francisco, CADublin, Ireland
Salary
$127,000–$249,000 / yr
Posted
101 days ago
Freshness
Confirmed live yesterday

Market check

Salary context

Competitive pay

How this pay compares to similar roles

Similar $181k
This role $188k
$112k most similar roles pay here $264k

This role pays more than 51% of similar roles. Most pay $147,200–$214,000 — the shaded band above. At the midpoint, this role pays about $188k versus about $181k for comparable roles.

Based on 240 similar postings.

Employer

About MongoDB

MongoDB is a leading American software company that develops and provides commercial support for a popular, source-available document database. Designed to handle unstructured and structured data natively, its platform is purpose-built for modern cloud applications, analytics, and AI experiences.

MongoDB currently has 311 open roles on FindRole.

Listed pay typically runs $126,000–$226,000 across 95 roles with salary data.

Most-posted roles

View all roles at MongoDB

At a glance

TL;DR · Senior Site Reliability Engineer, Fleet Management

As a Senior Site Reliability Engineer, Fleet Management, you will join the Platform Engineering team to manage the end-to-end lifecycle of a multi-cloud Kubernetes fleet. You will build and maintain a scalable, secure runtime environment that supports product needs across MongoDB while providing internal support for the Kubernetes ecosystem. Your daily work involves transitioning from Terraform-based infrastructure to an Operator-driven management model, resolving critical issues during on-call rotations, and performing blameless post-mortems to eliminate toil. The role requires proficiency in Go or Python, deep experience with containerization technologies like Kubernetes, and a strong grasp of Linux internals and networking concepts such as TCP/IP, DNS, and TLS. You will utilize tools including Helm, Kustomize, Gatekeeper, Kyverno, Crossplane, and ACK to manage infrastructure across AWS, GCP, and Azure platforms.

What does a Site Reliability Engineer earn in California?

Median $214000 from 54 postings across 15 companies.

See salary data

What you'll do

  • Develop and maintain a scalable, secure Kubernetes runtime environment to support product needs.
  • Provide internal technical support to engineering teams to resolve domain-specific Kubernetes issues.
  • Participate in a 24/7 on-call rotation to resolve critical infrastructure incidents.
  • Perform blameless post-mortems and implement systemic fixes to eliminate recurring production issues.
  • Automate manual operational processes to reduce toil and improve system reliability.
  • Manage the end-to-end lifecycle of the Kubernetes fleet, including CoreDNS, cert-manager, and Gatekeeper.
  • Transition infrastructure management from Terraform-based models to an Operator-driven lifecycle model.

What we're looking for

  • Have 6+ years of experience in software development and operating distributed systems.
  • Proficiency in Go, Python, or similar languages with a commitment to code quality and testing practices.
  • Deep experience using and extending containerization technologies, preferably Kubernetes.
  • Solid understanding of Linux operating system internals and networking concepts like TCP/IP, DNS, and TLS.
  • Strong operational ownership and a track record of debugging complex production issues.
  • Experience with Kubernetes ecosystem tools such as Helm, Kustomize, Gatekeeper, Kyverno, and CRDs/Operators.
  • Expertise in cloud infrastructure platforms including AWS, GCP, or Azure.
  • Proficiency in provisioning infrastructure using tools like Terraform, Crossplane, and AWS Controllers for Kubernetes (ACK).

More like this

Similar roles

Senior Site Reliability Engineer

MongoDB

Gurugram, India 52 days ago
Kubernetes Python Go AWS Google Cloud Platform Azure Linux TCP/IP DNS TLS Istio Cilium Service Mesh Distributed Systems Multi-cloud Alerting Networking
6+ yrs exp Hybrid

Security Software Engineer, Infrastructure Security

MongoDB

Remote (New York, NY) +3 101 days ago $127,000$249,000
Kubernetes eBPF Terraform AWS Azure GCP Python Golang Rust Java C/C++ Linux CI/CD Splunk Grafana Victoria Metrics AppArmor SELinux seccomp cgroups OPA Gatekeeper Kyverno Falco Tetragon
5+ yrs exp Remote

Senior Site Reliability Engineer

The Federal Reserve

Boston, MA 15 days ago $140,000$210,900
AWS EKS Terraform Python Java Go Docker Ansible CI/CD IaC Linux Shell Scripting Prometheus Grafana CloudWatch OpenSearch Dynatrace Consul Vault S3 RDS Aurora Route 53 ELB ECR

Senior Site Reliability Engineer

Autodesk

Remote (ID) +1 28 days ago $117,000$209,330
Site Reliability Engineering AWS Kubernetes Python Go Java Infrastructure as Code CI/CD CloudWatch Splunk Datadog Dynatrace Bash PowerShell FedRAMP Distributed Systems Load Balancing DNS
7+ yrs exp Remote

Senior Site Reliability Engineer

Autodesk

San Francisco, CA 28 days ago $117,000$209,330
SRE Python Go Java Bash PowerShell AWS Kubernetes Infrastructure as Code CI/CD CloudWatch Splunk Datadog Dynatrace FedRAMP Distributed Systems Load Balancing DNS
7+ yrs exp

Senior Site Reliability Engineer

Salesforce

Remote (San Francisco, CA) 58 days ago $148,500$223,900
SRE Python Go Docker Kubernetes CI/CD Prometheus Grafana ELK Splunk Datadog Temporal Airflow Argo Workflows AWS GCP Linux Unix LLM Prompt Engineering
5+ yrs exp Remote