Staff Site Reliability Engineer, Kubernetes

Okta Inc

Confirmed live 2 days ago High trust
Hybrid

Quick summary

Work type
Hybrid
Location
Bellevue, WAChicago, ILNew York, NYSan Francisco, CAWashington, DC
Salary
$194,000–$267,000 / yr
Posted
15 days ago
Freshness
Confirmed live 2 days ago

Market check

Salary context

Above market

How this pay compares to similar roles

Similar $183k
This role $230k
$134k most similar roles pay here $281k

This role pays more than 87% of similar roles. Most pay $151,678–$214,625 — the shaded band above. At the midpoint, this role pays about $230k versus about $183k for comparable roles.

Based on 240 similar postings.

Employer

About Okta Inc

Okta, Inc. is an American identity and access management company based in San Francisco. It provides cloud software that helps companies manage and secure user authentication into applications, and for developers to build identity controls into applications, websites, web services, and devices.[

Okta Inc currently has 137 open roles on FindRole.

Listed pay typically runs $174,000–$239,000 across 137 roles with salary data.

Most-posted roles

View all roles at Okta Inc

At a glance

TL;DR · Staff Site Reliability Engineer, Kubernetes

As a Staff Site Reliability Engineer - Kubernetes, you will join the engineering team to build and manage high-availability Kubernetes platforms supporting cloud-native applications. You will be responsible for architecting scalable infrastructure on AWS, managing EKS, ECS, S3, VPCs, and RDS while optimizing for cost and performance. Your daily work involves creating Helm charts, implementing Karpenter for dynamic scaling, and configuring Istio service mesh for traffic management and observability. To succeed, you must possess expertise in Terraform, CI/CD pipelines, and scripting languages like Python, Bash, or Go. You will also utilize monitoring tools such as Prometheus, Grafana, and the ELK Stack to ensure system reliability. This role focuses on solving complex infrastructure challenges by automating deployment workflows and maintaining secure, multi-region cloud environments for large-scale production workloads.

What does a Site Reliability Engineer earn in California?

Median $214000 from 54 postings across 15 companies.

See salary data

What you'll do

  • Design, implement, and maintain highly available and scalable Kubernetes platforms for production workloads.
  • Manage and optimize AWS cloud infrastructure including EKS, ECS, S3, VPCs, RDS, and IAM.
  • Create and manage Helm charts to automate the deployment of applications and services.
  • Implement and manage Karpenter to dynamically scale Kubernetes clusters based on demand.
  • Configure and manage Istio service mesh for traffic management, security, and observability.
  • Automate infrastructure and application deployments using CI/CD pipelines and scripting tools.
  • Respond to incidents and troubleshoot issues related to system performance, availability, and security.
  • Develop technical documentation for Kubernetes platform setup and operational procedures.

What we're looking for

  • Must have 4+ years of experience with Kubernetes and Helm.
  • Must have 4+ years of experience with Terraform.
  • Must have 5+ years of experience with AWS infrastructure.
  • Experience in multi-region cloud environments is required.
  • Proficiency in CI/CD pipelines, scripting (Python, Bash, or Go), and monitoring tools like Prometheus and Grafana is required.
  • Understanding of security best practices for cloud platforms and Kubernetes is preferred.
  • Bachelor’s degree in a related field or equivalent experience; CKA, CKAD, or AWS Certified DevOps Engineer certifications are preferred.
  • U.S. Citizen, National, Lawful Permanent Resident, Refugee, or Asylee.

More like this

Similar roles

Staff Site Reliability Engineer

Circle

Remote (San Francisco, CA) 37 days ago $195,000$257,500
Kubernetes Terraform Pulumi Go Python CI/CD Infrastructure as Code Blockchain Distributed Systems SQL Helm SRE Chaos Engineering Cloud Networking DNS
6+ yrs exp Remote

Staff Site Reliability Engineer

Okta Inc

Bellevue, WA +4 43 days ago $194,000$267,000
Splunk Terraform Go Python Ruby SPL Kubernetes EKS AWS GCP OpenTelemetry Vector Linux TCP/IP DNS Load Balancing Infrastructure as Code
5+ yrs exp Hybrid

Staff Site Reliability Engineer

TransUnion

Chicago, IL +4 140 days ago $112,500$187,500
GCP Kubernetes CI/CD Datadog Prometheus Grafana PagerDuty Linux PostgreSQL MySQL Cloud SQL Bigtable Firestore Redis Terraform Pulumi Python Bash Go Infrastructure-as-Code
5+ yrs exp Hybrid