Site Reliability Operations Engineer

Salesforce

Confirmed live today High trust

Quick summary

Work type
On-site
Location
Seattle, WA
Salary
$94,000–$142,300 / yr
Employment
Full-time
Posted
16 days ago
Freshness
Confirmed live today

Market check

Salary context

Below market

How this pay compares to similar roles

Similar $182k
This role $118k
$78k most similar roles pay here $241k

This role pays less than 93% of similar roles. Most pay $148,615–$215,000 — the shaded band above. At the midpoint, this role pays about $118k versus about $182k for comparable roles.

Based on 238 similar postings.

Employer

About Salesforce

Salesforce is the world''s leading customer relationship management (CRM) platform, offering cloud-based software for sales, service, marketing, analytics, and application development. Industry: Enterprise Software & Cloud Computing

Salesforce currently has 113 open roles on FindRole.

Listed pay typically runs $148,500–$223,900 across 106 roles with salary data.

Most-posted roles

View all roles at Salesforce

At a glance

TL;DR · Site Reliability Operations Engineer

As a Site Reliability Operations Engineer on the internal DET Site Reliability Operations team, you will support global employees by combining incident command, reliability engineering, and hands-on technical support. You will manage major incidents as an Incident Commander, coordinate technical teams for rapid service restoration, and monitor enterprise systems including infrastructure, applications, and network components. Your daily work involves troubleshooting across Windows and Linux servers, cloud platforms like AWS, and virtualization technologies while developing runbooks and automation to reduce toil. You will analyze KPI metrics to identify trends, perform root cause analyses, and manage problem management activities. Required skills include proficiency in ITIL frameworks, monitoring tools like Splunk or Grafana, and scripting in Python, Bash, or PowerShell. This role ensures business continuity by resolving complex technical issues across diverse platforms and vendor environments.

What you'll do

  • Manage major incidents affecting internal operations by serving as Incident Commander to coordinate technical teams and drive restoration.
  • Monitor and troubleshoot enterprise systems including infrastructure, applications, and network components across multiple platforms.
  • Create and improve runbooks, standard operating procedures, and automation to enhance global incident response.
  • Coordinate emergency changes and infrastructure updates to resolve critical issues and maintain business continuity.
  • Analyze incident data and KPI metrics to identify trends and provide actionable recommendations to stakeholders.
  • Lead problem management activities by investigating recurring incidents and documenting root cause analyses.
  • Participate in a 24/7 on-call rotation to handle escalations and serve as Duty Manager for high-severity events.
  • Identify and implement toil reduction opportunities to improve system reliability and operational efficiency.

What we're looking for

  • 5-8 years of experience in IT operations, incident management, or site reliability work.
  • A related technical degree is required.
  • Demonstrated ability to manage high severity incidents under pressure while balancing technical and business needs.
  • Strong verbal and written communication skills for explaining complex issues to technical and executive audiences.
  • Technical troubleshooting expertise across Windows/Linux servers, networking, cloud platforms, and virtualization technologies.
  • Experience with cloud platforms (e.g., AWS) and monitoring of IT infrastructure.
  • Understanding of the ITIL framework, specifically incident, problem, and change management processes.
  • Salesforce platform experience, industry certifications (ITIL, AWS, CCNA, MCSA, RHCE), scripting skills, or automation tool experience (preferred).

More like this

Similar roles

Site Reliability Operations Engineer

Salesforce

Seattle, WA 16 days ago $94,000–$142,300
Incident Management ITIL AWS Python Bash PowerShell Linux Splunk Grafana Tableau Puppet Chef Virtualization
5+ yrs exp

Site Reliability Engineer II

Mastercard

O Fallon, MO 11 days ago $76,000–$127,000
CI/CD DevOps Python Java Go C++ C Perl Ruby Git BitBucket Jenkins Maven Artifactory Chef Distributed Systems Scripting Monitoring Root Cause Analysis

Site Reliability Engineer

The Hartford

Hartford, CT +3 58 days ago $91,200–$136,800
SRE DevSecOps AWS Kubernetes Terraform CloudFormation Python Java CI/CD Splunk Dynatrace CloudWatch Oracle SQL Server Infrastructure as Code Agile LLM Microservices
3+ yrs exp Hybrid

Site Reliability Engineer

The Hartford

Hartford, CT +3 60 days ago $91,200–$136,800
SRE DevSecOps AWS Kubernetes Terraform CloudFormation Python Java Splunk Dynatrace CloudWatch CI/CD Infrastructure as Code Oracle SQL Server AI/ML Agile Microservices
3+ yrs exp Hybrid

Site Reliability Engineer

Berkeley Research Group

Remote 93 days ago $130,000–$160,000
Azure Kubernetes CI/CD GitHub Actions GitLab CI Golang Ruby Python AWS GCP Infrastructure as Code Datadog OpsGenie PagerDuty SRE Incident Management
5+ yrs exp Remote

Site Reliability Engineer

Balyasny Asset Management

Warsaw, Poland 94 days ago
Prometheus Grafana Loki Tempo OTEL Kubernetes Docker AWS Python Bash Go CI/CD DevOps SRE Agile
5+ yrs exp