Principal Site Reliability Engineer

Early Warning Services

Confirmed live yesterday High trust
Hybrid

Quick summary

Work type
Hybrid
Location
Scottsdale, AZChicago, ILSan Francisco, CANew York, NY
Salary
$194,000–$237,000 / yr
Posted
4 days ago
Freshness
Confirmed live yesterday

Market check

Salary context

Above market

How this pay compares to similar roles

Similar $183k
This role $216k
$133k most similar roles pay here $248k

This role pays more than 76% of similar roles. Most pay $150,537–$215,000 — the shaded band above. At the midpoint, this role pays about $216k versus about $183k for comparable roles.

Based on 238 similar postings.

Employer

About Early Warning Services

Early Warning Services is a fintech company that operates the Zelle person-to-person payments network, the Paze digital checkout wallet, and Certos fraud prevention and identity risk solutions for financial institutions.

Early Warning Services currently has 64 open roles on FindRole.

Listed pay typically runs $143,500–$183,000 across 60 roles with salary data.

Most-posted roles

View all roles at Early Warning Services

At a glance

TL;DR · Principal Site Reliability Engineer

Principal Site Reliability Engineer - Paze The Principal Site Reliability Engineer joins the team to apply software engineering and systems engineering practices to improve the reliability, resilience, scalability, and operational health of production services. This role involves partnering with Software Engineering teams to ensure observability, recoverability, performance, and operational readiness are integrated throughout the service lifecycle. You will build and improve CI/CD pipelines, Infrastructure as Code, automation, and monitoring tools while managing incident response and capacity management. The role requires expertise in distributed systems, Linux/Unix environments, and public cloud technologies like AWS, Azure, or GCP. Key technical requirements include proficiency in modern programming languages for scripting, experience with containers, and the ability to define SLIs and SLOs. This position addresses critical production challenges within a payment system infrastructure to ensure high availability and reduce operational toil through automated engineering practices.

What does a Site Reliability Engineer earn in California?

Median $214000 from 60 postings across 17 companies.

See salary data

What you'll do

  • Apply software engineering and automation principles to improve service reliability, scalability, and operational health.
  • Use data and rigorous analysis to identify reliability risks and guide technical decision-making.
  • Define and implement SLIs, SLOs, error budgets, and other metrics to monitor service health.
  • Improve observability through advanced logging, tracing, monitoring, alerting, and dashboarding.
  • Drive continuous improvement across CI/CD pipelines, Infrastructure as Code, and deployment practices.
  • Translate recurring production issues into improvements in code, architecture, and engineering tools.
  • Lead incident response efforts and conduct blameless post-incident learning to improve system resilience.
  • Reduce operational toil by creating reusable automation patterns and high-quality engineering practices.

What we're looking for

  • Candidates must independently possess the eligibility to work in the United States at the date of hire.
  • Minimum 15 years of relevant professional experience in SRE, Software Engineering, Systems Engineering, or related technical disciplines.
  • Experience with software development or scripting using one or more modern programming languages.
  • Experience with software engineering principles, distributed systems, production troubleshooting, automation, and observability.
  • Experience with public cloud technologies and architectures, preferably AWS.
  • Bachelor's degree in Computer Science, Software Engineering, Computer Engineering, Information Systems, or a related technical field, or equivalent practical experience.
  • Hands-on experience with AWS is preferred, or comparable experience with another major cloud platform (preferred).
  • Experience developing, deploying, operating, or improving highly available production software or distributed systems (preferred).

More like this

Similar roles

Staff Site Reliability Engineer

Early Warning Services

Scottsdale, AZ +3 4 days ago $131,000$160,000
AWS Azure GCP OCI Linux Unix CI/CD Infrastructure as Code Distributed Systems Site Reliability Engineering DevOps Monitoring Alerting Logging
8+ yrs exp Hybrid

Senior Site Reliability Engineer

Early Warning Services

Scottsdale, AZ +3 6 days ago
AWS CI/CD Infrastructure as Code Linux Unix Distributed Systems Monitoring Logging Tracing Networking Site Reliability Engineering
5+ yrs exp Hybrid

Principal Site Reliability Engineer

Oracle

Nashville, TN 61 days ago $84,900$209,500
Site Reliability Engineering Linux Windows Server Python Bash PowerShell Ansible Chef Oracle Cloud Infrastructure infrastructure-as-code Monitoring Logging Observability Capacity Planning incident-management Patching Citrix
3+ yrs exp

Lead Site Reliability Engineer

JPMorgan Chase

Plano, TX 48 days ago
Site Reliability Engineering Python Java Spring Boot .NET CI/CD Docker Kubernetes Terraform AWS Prometheus Grafana Dynatrace Datadog Splunk infrastructure-as-code CloudFormation ECS Incident Management Observability
5+ yrs exp

Principal Site Reliability Engineer

Early Warning Services

Scottsdale, AZ +3 118 days ago $194,000$237,000
Python Go Java Docker Kubernetes Microservices Kafka SQS JMS Oracle Dynamo DB Aurora Redis Memcached Linux CI/CD Git Chef Maven Jenkins AWS Ruby JavaScript
10+ yrs exp Hybrid

Principal Site Reliability Engineer

Oracle

Reston, VA +1 49 days ago $84,900$209,500
Kubernetes Terraform Docker Python Bash Linux Unix Oracle Database RAC Chef Puppet DNS DHCP HTTP TCP/IP LLM VMware Cisco
6+ yrs exp