Principal Site Reliability Engineer, Infrastructure Observability

T. Rowe Price

Confirmed live 2 days ago High trust
Hybrid

Quick summary

Work type
Hybrid
Location
Owings Mills, MD
Salary
$159,000–$272,000 / yr
Posted
176 days ago
Freshness
Confirmed live 2 days ago

Market check

Salary context

Above market

How this pay compares to similar roles

Similar $190k
This role $216k
$127k most similar roles pay here $288k

This role pays more than 75% of similar roles. Most pay $165,700–$215,000 — the shaded band above. At the midpoint, this role pays about $216k versus about $190k for comparable roles.

Based on 240 similar postings.

Employer

About T. Rowe Price

T. Rowe Price is an asset management firm focused on delivering global investment management excellence and retirement services

T. Rowe Price currently has 26 open roles on FindRole.

Listed pay typically runs $121,500–$207,500 across 26 roles with salary data.

Most-posted roles

View all roles at T. Rowe Price

At a glance

TL;DR · Principal Site Reliability Engineer, Infrastructure Observability

Principal Site Reliability Engineer, Infrastructure Observability will join a team focused on the observability, sustainability, scalability, measurability, and recoverability of cloud and on-prem solutions. This role involves formulating SRE strategies, designing technology solutions to minimize service disruptions, and fostering a culture of blameless post-mortems. The candidate will manage incident analysis for high-level trends, drive initiatives to reduce failures in complex distributed environments, and consolidate information from disconnected systems into cohesive views. Key technical requirements include experience with Amazon AWS, CI/CD toolchains, and the chaos model at scale. Proficiency is required in programming languages like Python, Java, Go, Node.js, or .Net Core, along with database development in SQL Server, PostgreSQL, or MySQL. The role utilizes tools such as New Relic, Elastic Stack, Prometheus, Grafana, Splunk, Ansible, Terraform, and Vault to ensure infrastructure reliability.

What does a Site Reliability Engineer earn?

Median $186200 from 131 postings across 36 companies.

See salary data

What you'll do

  • Design and implement technology solutions to prevent or minimize service disruptions across cloud and on-premise environments.
  • Develop automation and use best-of-breed tools to improve the observability, scalability, and recoverability of infrastructure.
  • Lead initiatives to reduce or prevent technology failures within complex, distributed systems.
  • Consolidate data from disconnected systems into cohesive views to identify trends, redundancies, and risks.
  • Drive the adoption of SRE standard methodologies and foster a culture of blameless post-mortems for continuous learning.
  • Define and report on Service Level Objectives (SLOs), Service Level Indicators (SLIs), and error budgets.
  • Standardize dashboards across the ecosystem for observability, application performance monitoring, and infrastructure logging.
  • Mentor team members and develop diverse talent within the Site Reliability Engineering function.

What we're looking for

  • Bachelor's degree or an equivalent combination of education and relevant experience.
  • 10+ years of experience designing and operating cloud infrastructure with senior-level impact.
  • 5+ years of experience building and supporting solutions in Amazon AWS.
  • 5+ years of experience building and running a DevOps and/or SRE function.
  • Fluency in multiple programming languages such as Python, Java, GO, Node.js, or .Net Core.
  • Proficiency with database development including SQL Server, PostgreSQL, and MySQL.
  • Experience with observability tools like New Relic, Elastic Stack, Prometheus, Grafana, and Splunk.
  • Experience with cloud management tools such as Ansible, Terraform, Vault, and Vagrant.

More like this

Similar roles

Principal Site Reliability Engineer

Nvidia

Santa Clara, CA 3 days ago $248,000$396,750
Kubernetes Distributed Systems Python Go Terraform AWS Azure GCP OpenTelemetry infrastructure-as-code Linux TypeScript JavaScript Java Crossplane AWS CDK CloudFormation AI/ML Platforms High-Performance Computing
10+ yrs exp Hybrid

Senior Site Reliability Engineer

Fiserv

Berkeley Heights, NJ 6 days ago $128,000$216,000
AWS Kubernetes Terraform CI/CD GitHub Actions Python Bash Ruby on Rails Docker Linux Unix RDBMS Document Storage New Relic Dynatrace Datadog DNS Load Balancing Virtual Networking

Senior Site Reliability Engineer

Adobe

New York, NY 35 days ago $177,900$257,550
AWS Kubernetes Python Terraform Docker CI/CD SageMaker Bedrock vLLM LangGraph PostgreSQL Aurora Memcached Ansible Chef Prometheus Grafana Splunk New Relic Fastly Node.js PHP Ruby Bash