Principal Software Engineer, SRE, Incident Response & Operational Excellence

Intuit

Confirmed live today High trust

Quick summary

Work type
On-site
Location
San Diego, CAMountain View, CA
Salary
$247,500–$335,000 / yr
Posted
2 days ago
Freshness
Confirmed live today

Market check

Salary context

Above market

How this pay compares to similar roles

Similar $209k
This role $291k
$120k $358k
below market most similar roles pay here above market

This role pays more than 86% of similar roles. Most pay $176,587–$241,550 — the blue band above. At the midpoint, this role pays about $291k versus about $209k for comparable roles.

Based on 240 similar postings.

Employer

About Intuit

Intuit is a financial software company known for products like TurboTax, QuickBooks, Mint, and Credit Karma, helping consumers and small businesses manage their finances and taxes. Industry: Financial Software & Technology

Intuit currently has 219 open roles on FindRole.

Listed pay typically runs $202,500–$274,000 across 198 roles with salary data.

Most-posted roles

View all roles at Intuit

At a glance

TL;DR · Principal Software Engineer, SRE, Incident Response & Operational Excellence

Principal Software Engineer – SRE, Incident Response & Operational Excellence serves as a Technical Duty Officer leading critical incidents for high-traffic production systems. Reporting to the Director of Cloud Engineering & Operations, this role involves commanding incident bridges, making decisive recovery calls, and directing communications during high-stakes outages. You will build and govern AI agents to accelerate triage and impact assessment while implementing engineering improvements to observability, runbooks, and escalation readiness. Key responsibilities include driving blameless post-incident reviews and defining company-wide standards for operational excellence. Required skills include proficiency in Python, Golang, or Java, Infrastructure as Code, and distributed systems across AWS, Kubernetes, and service meshes. You must demonstrate expertise in SLOs, metrics, logs, traces, and PagerDuty within a financial services context to ensure rapid recovery of customer workflows.

What does a Software Engineer earn in California?

Median $214500 from 1223 postings across 78 companies.

See salary data

What you'll do

  • Lead critical incidents as the Technical Duty Officer with ultimate authority over recovery actions and communications.
  • Direct fast recovery by making decisive calls on rollbacks, failovers, traffic shifts, and feature disablement.
  • Partner with business units to define and instrument real-time customer impact signals for high-value workflows.
  • Improve observability by partnering with platform teams on SLOs, dashboards, alerting, and tracing.
  • Own the clarity of runbooks and ensure recurring validation of notification channels and escalation policies.
  • Drive blameless post-incident reviews to identify systemic remediation and track high-leverage actions to closure.
  • Develop and orchestrate AI agents to automate triage, impact assessment, and post-incident learning.
  • Provide 24x7 coverage in a follow-the-sun rotation, including off-hours and peak-season readiness.

What we're looking for

  • BS in Computer Science or equivalent work-related experience.
  • 12+ years of engineering experience, including significant time operating large-scale, high-traffic production systems.
  • Proven track record commanding large-scale incidents in an enterprise-level or financial-services environment.
  • Deep operational excellence and SRE expertise, including SLOs, observability, alerting design, and formal incident-management frameworks.
  • Broad, hands-on distributed-systems intuition across cloud infrastructure (AWS), Kubernetes, networking, data stores, and messaging.
  • Proficiency in scripting and development (e.g., Python, Golang, Java), Infrastructure as Code, and automation.
  • Fluent in AI agent development and orchestration to build agentic and self-service tooling for triage and post-incident analysis.
  • Willingness to participate in a 24x7 follow-the-sun on-call rotation, including off-hours and peak-season coverage.

More like this

Similar roles

Senior Engineer, Incident Response Engineering

Target

Brooklyn Park, MN 167 days ago
SOAR Python TypeScript JavaScript React REST APIs Data Pipelines Backend Development Automation Workflows Incident Response Testing Debugging
5+ yrs exp Hybrid

Principal Software Engineer

Autodesk

Atlanta, GA 58 days ago
Distributed Systems Data Pipelines NoSQL Object Storage Serverless API Design ETL DevOps Search Relevance Schema Validation
8+ yrs exp Hybrid

Principal Software Engineer

General Dynamics

San Antonio, TX 84 days ago $119,000–$161,000
DevSecOps CI/CD Kubernetes Python Java Terraform Ansible Docker AWS Azure Google Cloud Go Bash infrastructure-as-code NIST OWASP Splunk TCP/IP DNS HTTP TLS Agile Methodology
10+ yrs exp Hybrid

Principal Software Engineer

Microsoft

79 days ago $142,800–$274,800
Distributed Databases Storage Systems C C++ C# Java Python JavaScript Cloud Infrastructure Service Architecture Capacity Planning Data Protection Telemetry Root-Cause Analysis Operational Automation
6+ yrs exp Hybrid

Principal Software Engineer

Twilio

Remote 80 days ago $188,240–$235,300
AWS Terraform Helm Clickhouse Kafka Spark Python Java Go Bash Distributed Systems High Availability Observability Network Security
10+ yrs exp Remote

Principal Software Engineer

Nordstrom

Seattle, WA 72 days ago $191,000–$297,000
Java Python AWS GCP Kafka GitLab CI/CD Generative AI Agentic AI AWS Bedrock Microservices Big Data RFID WMS
10+ yrs exp