Manager, Site Reliability Engineering

Okta Inc

Confirmed live yesterday High trust

Quick summary

Work type
On-site
Location
New York, NYWashington, DC
Salary
$182,000–$250,800 / yr
Posted
16 days ago
Freshness
Confirmed live yesterday

Market check

Salary context

Above market

How this pay compares to similar roles

Similar $186k
This role $216k
$135k most similar roles pay here $263k

This role pays more than 76% of similar roles. Most pay $156,778–$215,875 — the shaded band above. At the midpoint, this role pays about $216k versus about $186k for comparable roles.

Based on 238 similar postings.

Employer

About Okta Inc

Okta, Inc. is an American identity and access management company based in San Francisco. It provides cloud software that helps companies manage and secure user authentication into applications, and for developers to build identity controls into applications, websites, web services, and devices.[

Okta Inc currently has 137 open roles on FindRole.

Listed pay typically runs $174,000–$239,000 across 137 roles with salary data.

Most-posted roles

View all roles at Okta Inc

At a glance

TL;DR · Manager, Site Reliability Engineering

Manager, Site Reliability Engineering (Auth0) joins the SRE Leadership Team to ensure the reliability and operational excellence of the Auth0 platform. This role involves leading technical direction, translating organizational vision into actionable roadmaps, and managing cross-functional initiatives across product and platform teams. Day-to-day responsibilities include participating in 24/7 on-call rotations to troubleshoot critical systems, building infrastructure resilience through monitoring and automation, and mentoring SRE talent via pair programming and code reviews. The ideal candidate possesses deep expertise in cloud platforms like AWS and Azure, utilizes Terraform for infrastructure as code, and demonstrates strong programming skills in Go or Python. This role addresses the challenge of maintaining a trusted authentication platform for millions of users by embedding observability, resilience, and software engineering rigor into all technical decisions to ensure high availability at scale.

What you'll do

  • Translate organizational vision into actionable technical roadmaps and lead cross-functional initiatives across product and platform teams.
  • Participate in 24/7 on-call rotations to troubleshoot and remediate incidents on critical systems.
  • Design and implement monitoring, alerting, and automation improvements to reduce toil and improve infrastructure resilience.
  • Establish policies and cultural standards for observability, reliability, and software engineering rigor across all engineering efforts.
  • Mentor SRE talent through pair programming, design discussions, and code reviews to foster professional growth.
  • Represent the reliability team in architectural reviews and strategic planning to ensure reliability is a core consideration in major decisions.

What we're looking for

  • You must have at least 3 years of hands-on team leadership in SRE or software engineering roles in cloud-native environments.
  • You must have at least 8 years of total industry experience.
  • You must possess deep expertise in cloud platforms (AWS, Azure) and infrastructure as code using Terraform.
  • You must have proven experience managing cloud-native architectures including containers, Kubernetes, microservices, and databases.
  • You must have strong programming skills in Go or Python to build production-grade tools and automation.
  • You must possess excellent verbal and written communication skills for high-pressure incident management and stakeholder engagement.
  • You must be able to provide documentation of U.S. Person status (e.g., U.S. Citizen, National, Lawful Permanent Resident, Refugee, or Asylee).
  • Experience with open-source infrastructure, observability tooling, or automated runbook systems is preferred.

More like this

Similar roles

Manager, Site Reliability Engineering

Okta Inc

San Francisco, CA 43 days ago $204,000$306,000
AWS Kubernetes Terraform CI/CD Grafana Splunk APM DevOps SaaS Cloud-native Architecture Containerization Edge Networking SDLC
3+ yrs exp Hybrid

Senior Manager, Site Reliability Engineering

Okta Inc

Washington, DC 43 days ago $207,000$284,900
SRE Kubernetes Terraform CI/CD AWS GovCloud Incident Management Observability NIST SP 800-53 FedRAMP FISMA DISA STIGs Automation Runbooks
5+ yrs exp Hybrid

Senior Site Reliability Engineer

Okta Inc

Bellevue, WA +1 18 days ago $147,000$202,400
Terraform Kubernetes Spinnaker Flyway Snowflake CI/CD Infrastructure as Code Containerization SaaS AIOps Cloud Infrastructure Automation
Hybrid

Principal Site Reliability Engineer

Nvidia

Santa Clara, CA 3 days ago $248,000$396,750
Kubernetes Distributed Systems Python Go Terraform AWS Azure GCP OpenTelemetry infrastructure-as-code Linux TypeScript JavaScript Java Crossplane AWS CDK CloudFormation AI/ML Platforms High-Performance Computing
10+ yrs exp Hybrid

Manager, Site Reliability Engineering

Oracle

Reston, VA +1 49 days ago $102,000$234,600
Site Reliability Engineering Incident Response Automation Monitoring Provisioning Decommissioning Scalability Data Collection
6+ yrs exp

Senior Manager, Site Reliability Engineering

Oracle

Nashville, TN 30 days ago $121,500$264,100
AIOps Automation Business Intelligence Data Center Operations Incident Management Network Operations Network Routing Structured Cabling Deployment Telecommunications Networking
3+ yrs exp