Staff Site Reliability Engineer

Okta Inc

Confirmed live today High trust
Hybrid

Quick summary

Work type
Hybrid
Location
Bellevue, WAChicago, ILNew York, NYSan Francisco, CAWashington, DC
Salary
$194,000–$267,000 / yr
Posted
66 days ago
Freshness
Confirmed live today

Market check

Salary context

Above market

How this pay compares to similar roles

Similar $185k
This role $230k
$130k most similar roles pay here $282k

This role pays more than 85% of similar roles. Most pay $151,612–$217,500 — the shaded band above. At the midpoint, this role pays about $230k versus about $185k for comparable roles.

Based on 240 similar postings.

Employer

About Okta Inc

Okta, Inc. is an American identity and access management company based in San Francisco. It provides cloud software that helps companies manage and secure user authentication into applications, and for developers to build identity controls into applications, websites, web services, and devices.[

Okta Inc currently has 105 open roles on FindRole.

Listed pay typically runs $174,000–$239,000 across 105 roles with salary data.

Most-posted roles

View all roles at Okta Inc

At a glance

TL;DR · Staff Site Reliability Engineer

Staff Site Reliability Engineer - Splunk joins the team to own and evolve the company's Splunk ecosystem into a comprehensive, scalable Observability Platform. This role involves treating infrastructure as code to automate the deployment of agents and collectors across complex distributed systems while optimizing the collection, processing, and storage of log data for high reliability and low latency. The successful candidate will eliminate toil through automation, participate in on-call rotations, and lead post-incident reviews to drive observability-driven development. Key technical requirements include expertise in Splunk Cloud at scale, including Workload Management and HEC optimization, along with proficiency in SPL, Go, Python, or Ruby. The role requires deep knowledge of Linux internals, networking protocols like TCP/IP and DNS, and container orchestration using Kubernetes or EKS to solve complex cross-service performance bottlenecks.

What does a Site Reliability Engineer earn in California?

Median $214000 from 61 postings across 17 companies.

See salary data

What you'll do

  • Design and maintain scalable observability infrastructure using Terraform as code.
  • Optimize the collection, processing, and storage of log data within the Splunk ecosystem.
  • Automate the deployment and scaling of observability agents and collectors to eliminate manual toil.
  • Create intuitive and actionable Splunk dashboards that correlate data across multiple sources.
  • Participate in on-call rotations and lead post-incident reviews to drive systemic improvements.
  • Develop internal tools and automate workflows using Go, Python, or Ruby.
  • Troubleshoot complex cross-service performance bottlenecks using a data-driven approach.

What we're looking for

  • Minimum 5 years of experience scaling and managing Splunk Cloud at scale (1000+ SVCs).
  • Minimum 5 years of experience in an SRE, DevOps, or Systems Engineering role focused on high-availability systems.
  • Strong coding skills in SPL, Go, Python, or Ruby to build tools and automate workflows.
  • Proficiency in infrastructure as code using Terraform.
  • Deep understanding of Linux internals, networking (TCP/IP, DNS, Load Balancing), and Kubernetes/EKS.
  • Experience with OpenTelemetry, Vector, or similar frameworks for instrumenting applications (preferred).
  • Experience implementing Splunk charge-back app for usage reporting (preferred).
  • U.S. Citizen, National, Lawful Permanent Resident, Refugee, or Asylee.

More like this

Similar roles

Staff Site Reliability Engineer

Okta Inc

Bellevue, WA +4 66 days ago $194,000–$267,000
Splunk Terraform Go Python Ruby SPL Kubernetes EKS AWS GCP OpenTelemetry Vector Linux TCP/IP DNS Load Balancing Infrastructure as Code
5+ yrs exp Hybrid

Senior Staff Site Reliability Engineer

Cisco

Remote (Irvine, CA) +1 38 days ago $192,400–$275,800
SRE Linux Administration AWS GCP Azure Python Go Distributed Systems Splunk SPL Indexer Clustering Search Head Clusters KVStore Monitoring Alerting Observability Root Cause Analysis
10+ yrs exp Remote

Staff Site Reliability Engineer, Kubernetes

Okta Inc

Bellevue, WA +4 38 days ago $194,000–$267,000
Kubernetes AWS Helm Karpenter Istio Terraform CI/CD Python Bash Go Docker Prometheus Grafana CloudWatch ELK Stack S3 RDS EC2 IAM
Hybrid

Staff Site Reliability Engineer, Kubernetes

Okta Inc

Bellevue, WA +4 38 days ago $194,000–$267,000
Kubernetes AWS Terraform Helm Istio Karpenter CI/CD Python Go Bash Docker Prometheus Grafana CloudWatch ELK Stack S3 RDS EC2 IAM CloudFormation
Hybrid

Staff Site Reliability Engineer

CME Group

Chicago, IL 65 days ago $132,100–$220,100
Python Go Kubernetes GCP GKE Kafka Terraform ArgoCD Node.js Gemini Distributed Systems GitOps SRE
10+ yrs exp Hybrid

Staff Site Reliability Engineer

Circle

Remote (San Francisco, CA) 60 days ago $195,000–$257,500
Kubernetes Terraform Pulumi Go Python CI/CD Infrastructure as Code Blockchain Distributed Systems SQL Helm SRE Chaos Engineering Cloud Networking DNS
6+ yrs exp Remote