Senior Staff Site Reliability Engineer, Storage Layer Services

MongoDB

Confirmed live yesterday Trusted
Remote

Quick summary

Work type
Remote
Location
Toronto, CanadaMontreal, Canada
Salary
$144,000–$200,000 / yr
Posted
101 days ago
Freshness
Confirmed live yesterday

Market check

Salary context

Below market

How this pay compares to similar roles

Similar $188k
This role $172k
$134k most similar roles pay here $240k

This role pays less than 67% of similar roles. Most pay $159,625–$217,225 — the shaded band above. At the midpoint, this role pays about $172k versus about $188k for comparable roles.

Based on 238 similar postings.

Employer

About MongoDB

MongoDB is a leading American software company that develops and provides commercial support for a popular, source-available document database. Designed to handle unstructured and structured data natively, its platform is purpose-built for modern cloud applications, analytics, and AI experiences.

MongoDB currently has 311 open roles on FindRole.

Listed pay typically runs $126,000–$226,000 across 95 roles with salary data.

Most-posted roles

View all roles at MongoDB

At a glance

TL;DR · Senior Staff Site Reliability Engineer, Storage Layer Services

Site Reliability Engineer (Senior or Staff), Storage Layer Services (SLS) joins the Storage Layer Services team to help re-architect and manage the multi-tenant distributed storage services for the Atlas platform. This role involves defining SLOs, shaping capacity plans, and ensuring the reliability, durability, and operational safety of the underlying cloud storage infrastructure. You will build resilient, self-healing systems while identifying key metrics to monitor service health and performance from the application level down to the kernel. The work focuses on solving complex problems related to multi-tenant distributed storage and large-scale data migrations. Required skills include proficiency in Python or Go, experience with Kubernetes and containerization technologies, and expertise in cloud platforms like AWS, GCP, or Azure. Candidates must also possess a deep understanding of Linux operating system internals, networking concepts such as TCP/IP and DNS, and stateful database systems.

What does a Site Reliability Engineer earn?

Median $186200 from 131 postings across 36 companies.

See salary data

What you'll do

  • Build and maintain multi-tenant distributed storage systems while balancing long-term infrastructure goals with immediate engineering needs.
  • Ensure the reliability, durability, and fault tolerance of services and underlying infrastructure.
  • Identify and configure key metrics to detect incidents and quantify service health and performance.
  • Participate in a 24/7 on-call rotation to resolve issues involving storage infrastructure.
  • Optimize infrastructure performance from the application level down to the kernel.
  • Automate manual processes to reduce toil and improve operational efficiency.
  • Manage and scale infrastructure across multi-cloud environments including AWS, GCP, and Azure.

What we're looking for

  • Have 6+ years of experience working on software development and operating distributed systems.
  • Proficiency in Python, Go, or a similar programming language.
  • Experience operating or supporting stateful storage or database systems at scale.
  • Experience using and extending containerization technologies, particularly Kubernetes.
  • Expertise in cloud infrastructure platforms including AWS, Google Cloud Platform (GCP), or Azure.
  • Understanding of Linux operating system internals and networking concepts like TCP/IP, DNS, TLS, and routing.
  • Ability to participate in a 24/7 on-call rotation to resolve infrastructure issues.

More like this

Similar roles

Site Reliability Engineer

MongoDB

Remote (New York, NY) 101 days ago $127,000$249,000
SRE AWS Azure GCP Linux Go Ruby Python HTTP TLS DNS Multi-cloud Automation
5+ yrs exp Remote Hybrid