Principal Site Reliability Engineer

Oracle

Confirmed live today High trust

Quick summary

Work type
On-site
Location
NCNashville, TN
Salary
$84,900–$209,500 / yr
Posted
3 days ago
Freshness
Confirmed live today

Market check

Salary context

Below market

How this pay compares to similar roles

Similar $181k
This role $147k
$68k most similar roles pay here $242k

This role pays less than 75% of similar roles. Most pay $147,355–$215,000 — the shaded band above. At the midpoint, this role pays about $147k versus about $181k for comparable roles.

Based on 238 similar postings.

Employer

About Oracle

Oracle Corporation is a leading multinational technology company specializing in database software, cloud computing, and enterprise software.

Oracle currently has 202 open roles on FindRole.

Listed pay typically runs $102,300–$209,500 across 184 roles with salary data.

Most-posted roles

View all roles at Oracle

At a glance

TL;DR · Principal Site Reliability Engineer

As a Principal Site Reliability Engineer, you will join the team to design and architect infrastructure and services ensuring reliability and functionality. You will be responsible for forecasting demand, managing capacity needs, and collaborating with software development teams to build scalable infrastructures. Your daily work involves performing incident response, root cause analysis, and maintenance tasks like software installs and security updates. You will identify automation opportunities, develop scripts to mitigate defects, and conduct advanced experiments with new tools to optimize performance. The role requires expertise in data collection, triage, and technical analysis to maintain service level objectives. You will also provide comprehensive health reporting, manage on-call shifts, and communicate complex information regarding scale, security, and performance attributes while proactively identifying improvements for infrastructure bottlenecks and deployment efficiency across various services.

What does a Site Reliability Engineer earn?

Median $188000 from 130 postings across 40 companies.

See salary data

What you'll do

  • Design and architect infrastructure and services to ensure reliability, scalability, and functionality.
  • Forecast demand and manage capacity to ensure systems have sufficient resources for current and future workloads.
  • Perform incident response, root cause analysis, and maintenance tasks including software updates and recovery.
  • Develop and implement automation tools and scripts to improve operational efficiency and mitigate defects.
  • Provide comprehensive health reporting and proactively communicate the impact of infrastructure or tool changes.
  • Conduct advanced experiments with new technologies to optimize performance and stay ahead of site reliability trends.
  • Serve as a key escalation point for complex technical issues while participating in on-call shifts.
  • Document incidents, perform post-mortem procedures, and provide data to drive business development decisions.

What we're looking for

  • Applicants must be able to read, write, and speak English.
  • Candidates must have 3 to 5+ years of experience in site reliability or related fields.
  • Experience designing and architecting infrastructure for reliability and functionality is required.
  • Ability to perform incident response, root cause analysis, and technical troubleshooting is required.
  • Proficiency in developing automation tools and scripts to improve operational efficiency is required.
  • Ability to provide comprehensive health reporting and communicate complex technical information to stakeholders is required.
  • Experience managing moderately complex projects and coordinating tasks across multiple teams is required.
  • Ability to mentor junior team members and participate in the candidate interview process is required.

More like this

Similar roles

Principal Site Reliability Engineer

Early Warning Services

San Francisco, CA +3 4 days ago $173,000$230,000
Site Reliability Engineering DevOps AWS Infrastructure as Code CI/CD Linux Unix Distributed Systems Monitoring Logging Tracing Incident Management Capacity Management
10+ yrs exp Hybrid

Principal Site Reliability Engineer

Early Warning Services

Scottsdale, AZ +3 124 days ago $194,000$237,000
Python Go Java Docker Kubernetes Microservices Kafka SQS JMS Oracle Dynamo DB Aurora Redis Memcached Linux CI/CD Git Chef Maven Jenkins AWS Ruby JavaScript
10+ yrs exp Hybrid

Senior Site Reliability Engineer

Oracle

Nashville, TN 3 days ago $81,100$187,000
Automation Incident Response Monitoring Data Collection Provisioning Deployment Scalability
3+ yrs exp

Principal Site Reliability Engineer

Nvidia

Santa Clara, CA 13 days ago $248,000$396,750
Kubernetes Distributed Systems Python Go Terraform AWS Azure GCP OpenTelemetry infrastructure-as-code Linux TypeScript JavaScript Java Crossplane AWS CDK CloudFormation AI/ML Platforms High-Performance Computing
10+ yrs exp Hybrid