Senior Site Reliability Engineer

Salesforce

Confirmed live yesterday High trust
Remote

Quick summary

Work type
Remote
Location
San Francisco, CA
Salary
$148,500–$223,900 / yr
Posted
58 days ago
Freshness
Confirmed live yesterday

Market check

Salary context

Competitive pay

How this pay compares to similar roles

Similar $182k
This role $186k
$133k most similar roles pay here $235k

This role pays less than 52% of similar roles. Most pay $150,000–$213,125 — the shaded band above. At the midpoint, this role pays about $186k versus about $182k for comparable roles.

Based on 238 similar postings.

Employer

About Salesforce

Salesforce is the world''s leading customer relationship management (CRM) platform, offering cloud-based software for sales, service, marketing, analytics, and application development. Industry: Enterprise Software & Cloud Computing

Salesforce currently has 106 open roles on FindRole.

Listed pay typically runs $148,500–$260,100 across 98 roles with salary data.

Most-posted roles

View all roles at Salesforce

At a glance

TL;DR · Senior Site Reliability Engineer

Senior Site Reliability Engineer The Senior Site Reliability Engineer joins the Site Reliability organization to ensure cloud service availability and performance. This role involves leading incident detection, response, and resolution while proactively designing systems that utilize automation, observability, and AI-powered platforms to reduce toil. The engineer will build production-grade observability solutions, implement self-healing systems, and develop AI-driven operational tools like anomaly detection and predictive analysis pipelines. Key technical requirements include proficiency in Python and Go, experience with containerized architectures using Docker and Kubernetes, and knowledge of distributed systems including DNS, HTTP, and load balancing. The role also requires expertise in workflow engines such as Temporal, Airflow, or Argo Workflows. Working within a follow-the-sun model, the engineer ensures high availability for global services by managing SLIs, SLOs, and error budgets while mentoring junior team members through code reviews and technical coaching.

What does a Site Reliability Engineer earn in California?

Median $214000 from 54 postings across 15 companies.

See salary data

What you'll do

  • Lead incident detection, response, and resolution while conducting root cause analysis and blameless postmortems.
  • Design and implement automated, self-healing systems using workflow engines like Temporal, Airflow, or Argo Workflows.
  • Build production-grade observability solutions including monitoring, logging, alerting, and tracing to ensure high availability.
  • Develop AI/ML-powered operational tools for anomaly detection, predictive analysis, and intelligent runbook automation.
  • Partner with product teams to define and maintain SLIs, SLOs, and error budgets for service reliability.
  • Automate manual workflows using AI-driven insights to reduce operational toil and improve system performance.
  • Provide technical coaching and mentorship to junior engineers through code reviews and pair programming.
  • Develop high-quality software in Python and Go while integrating AI tools into the development workflow.

What we're looking for

  • A related technical degree is required.
  • 5+ years of experience in systems engineering and software engineering for large-scale, internet-facing services.
  • Proficiency in Python and Go (GoLang) with strong software engineering practices like testing and CI/CD.
  • Hands-on expertise with containerized architectures (Docker, Kubernetes) and distributed systems.
  • Experience building and operating observability platforms using tools like Grafana, Prometheus, or Datadog.
  • Proven experience in incident management, including on-call participation, root cause analysis, and postmortems.
  • Experience applying AI/ML to operations, including anomaly detection, predictive analysis, and prompt engineering.
  • Proficiency with workflow orchestration engines such as Temporal, Airflow, or Argo Workflows.

More like this

Similar roles

Senior Site Reliability Engineer

Fiserv

Sunnyvale, CA 12 days ago $160,000$240,000
GCP Kubernetes Terraform Ansible Puppet Prometheus Grafana Datadog HAProxy GitHub Actions Python Go Java Shell Scripting Infrastructure as Code SLIs SLOs

Senior Site Reliability Engineer

Okta Inc

Bellevue, WA +1 18 days ago $147,000$202,400
Terraform Kubernetes Spinnaker Flyway Snowflake CI/CD Infrastructure as Code Containerization SaaS AIOps Cloud Infrastructure Automation
Hybrid

Senior Lead Site Reliability Engineer

JPMorgan Chase

Palo Alto, CA 59 days ago
Site Reliability Engineering Java Go Python Terraform Kubernetes Docker CI/CD GitOps Grafana Prometheus Dynatrace Datadog Splunk Kafka RabbitMQ SQS Neo4j Pinecone Weaviate Chroma LangChain LangGraph AutoGen CrewAI GitHub Copilot Fluentd Logstash Vector RESTful APIs RAG TensorFlow PyTorch scikit-learn Hadoop Spark Flink MongoDB Cassandra DynamoDB InfluxDB TimescaleDB AWS Azure GCP
5+ yrs exp

Senior Site Reliability Engineer

The Federal Reserve

Boston, MA 15 days ago $140,000$210,900
AWS EKS Terraform Python Java Go Docker Ansible CI/CD IaC Linux Shell Scripting Prometheus Grafana CloudWatch OpenSearch Dynatrace Consul Vault S3 RDS Aurora Route 53 ELB ECR

Senior Site Reliability Engineer

Autodesk

Remote (ID) +1 28 days ago $117,000$209,330
Site Reliability Engineering AWS Kubernetes Python Go Java Infrastructure as Code CI/CD CloudWatch Splunk Datadog Dynatrace Bash PowerShell FedRAMP Distributed Systems Load Balancing DNS
7+ yrs exp Remote

Senior Site Reliability Engineer

Autodesk

San Francisco, CA 28 days ago $117,000$209,330
SRE Python Go Java Bash PowerShell AWS Kubernetes Infrastructure as Code CI/CD CloudWatch Splunk Datadog Dynatrace FedRAMP Distributed Systems Load Balancing DNS
7+ yrs exp