Quick summary
- Work type
- Remote
- Location
- Remote
- Salary
- $217,000–$303,900 / yr
- Posted
- 28 days ago
- Freshness
- Confirmed live yesterday
Market check
Salary context
How this pay compares to similar roles
This role pays more than 88% of similar roles. Most pay $159,375–$235,800 — the shaded band above. At the midpoint, this role pays about $260k versus about $198k for comparable roles.
Based on 240 similar postings.
Employer
About Reddit
Reddit is a social news aggregation and discussion platform where users share content, vote on posts, and engage in community conversations across thousands of interest-based forums called subreddits.
Reddit currently has 77 open roles on FindRole.
Listed pay typically runs $217,000–$303,400 across 77 roles with salary data.
Most-posted roles
- Software Engineer 19
- Data Scientist 8
- Machine Learning Engineer 5
- Frontend Engineer 4
- Machine Learning Systems Engineer 4
At a glance
TL;DR · Staff Software Engineer, Observability
Staff Software Engineer, Observability joins the Observability team to work at the intersection of infrastructure and software development. This role involves building and maintaining foundational platforms for infrastructure, focusing on improving availability, scalability, latency, and efficiency across monitoring, logging, and distributed tracing systems. The engineer will contribute to technical strategy, automate event-driven processes, manage on-call responsibilities, and submit upstream changes to open-source projects. Key technologies include Prometheus, Thanos, Grafana, Vector, Clickhouse, Loki, and OTEL. Candidates should possess experience in distributed systems, Kubernetes development, and troubleshooting complex software environments at scale. The role addresses the technical challenges of managing massive data volumes, including performance engineering on distributed query systems and developing new ways for users to interpret telemetry data within a high-scale infrastructure environment.
What does a Software Engineer earn in Remote?
Median $204500 from 415 postings across 57 companies.
Skills
What you'll do
- Develop software to improve the availability, scalability, latency, and efficiency of observability components.
- Maintain and scale large-scale monitoring systems using Prometheus and Thanos.
- Manage and optimize a high-volume logging platform built on Vector and Clickhouse.
- Build new features and tools for distributed tracing using OTEL, Clickhouse, and Grafana.
- Automate critical aspects of the event-driven development process.
- Contribute upstream changes to open-source projects used by the infrastructure team.
- Participate in on-call rotations to troubleshoot system and software issues.
What we're looking for
- 7+ years of experience developing internet-scale software, preferably in the context of infrastructure.
- Experience developing on top of Kubernetes or similar distributed systems.
- Strong troubleshooting capabilities surrounding both systems and software.
- Experience engineering large systems, tracking work, and being a self-starter on projects.
- Excellent communication skills to collaborate with a service-oriented team and company.
- Familiarity with distributed systems development (preferred).
- Familiarity with specific tools such as Prometheus, Thanos, Grafana, Vector, Clickhouse, Otel, or Loki (preferred).
- Kubernetes controller or operator development experience (preferred).
More like this
Similar roles
Staff Site Reliability Engineer, Ads
Staff Software Engineer, Observability
Robinhood
Staff Software Engineer
PayPal
Staff Software Engineer
Datadog
Staff Software Engineer
GEICO