Principal System Engineering, SRE

AT&T

Confirmed live yesterday High trust
Closes in 7 days

Quick summary

Work type
On-site
Location
Dallas, TXPlano, TX
Salary
$155,400–$261,100 / yr
Posted
8 days ago
Freshness
Confirmed live yesterday
Closes
Sep 18, 2026 (soon)

Market check

Salary context

Competitive pay

How this pay compares to similar roles

Similar $185k
This role $208k
$129k most similar roles pay here $275k

This role pays more than 58% of similar roles. Most pay $156,600–$214,125 — the shaded band above. At the midpoint, this role pays about $208k versus about $185k for comparable roles.

Based on 240 similar postings.

Employer

About AT&T

AT&T is a US-based telecommunications company providing wireless, broadband, and fiber internet service along with phone and connectivity products for consumers and businesses.

AT&T currently has 61 open roles on FindRole.

Listed pay typically runs $117,350–$226,800 across 56 roles with salary data.

Most-posted roles

View all roles at AT&T

At a glance

TL;DR · Principal System Engineering, SRE

As a Principal System Engineering - SRE within the Systems Reliability and Software Delivery team, you will focus on identifying why production incidents occur and implementing long-term preventive measures. You will analyze end-to-end system architecture across cloud environments, web applications, APIs, and databases to perform deep root cause analysis. By utilizing observability data, you will create structured postmortems and partner with engineering teams to drive corrective actions. The role requires expertise in tools such as T-APM, T-Trace, CatchPoint, Grafana, and ServiceNow, alongside proficiency in Python, SQL, Power BI, and Tableau. You will leverage AI-assisted analysis and data analytics to move the organization from reactive responses to proactive reliability. Key technical competencies include experience with distributed systems, CI/CD pipelines, Agile methodologies, and modern enterprise release management within complex infrastructure environments.

What you'll do

  • Analyze production incidents end-to-end across applications, infrastructure, and cloud environments.
  • Use observability data to identify root causes, patterns, and systemic weaknesses.
  • Create high-quality postmortems based on incident insights.
  • Partner with engineering teams to implement permanent fixes and preventive improvements.
  • Utilize automation and AI-assisted analysis to transition from reactive response to proactive reliability.
  • Analyze operational data using tools, queries, or advanced analytics for pattern detection.

What we're looking for

  • Must have 7+ years of experience in Systems Engineering, ITSM, or Release/Change Management.
  • Must have a professional background in SRE, Support, or QA.
  • Must possess hands-on experience with observability tools such as T-APM, T-Trace, CatchPoint, or Grafana.
  • Must be proficient in Python, SQL, Power BI, and ITSM tools like ServiceNow.
  • Must have experience with AI technologies, data analytics, and building Gen AI use cases.
  • Must understand modern enterprise Release Management/Change Management for Agile and DevOps environments.
  • A Bachelor’s degree in Computer Science is preferred.
  • Relevant certifications in SAFe, Agile, DevOps, or AI/ML are preferred.

More like this

Similar roles

Principal System Engineering, SRE

AT&T

Plano, TX 8 days ago $155,400$261,100
SRE Python SQL Power BI Tableau Grafana CatchPoint ServiceNow Jira Cloud Git CI/CD Agile SAFe DevOps Data Analytics Gen AI ITSM Release Management Change Management Distributed Systems
7+ yrs exp

Principal System Engineering, SRE

AT&T

Atlanta, GA 8 days ago $155,400$261,100
SRE Python SQL Power BI Tableau Grafana CatchPoint ServiceNow Jira Cloud Git CI/CD Agile SAFe DevOps Data Analytics Gen AI ITSM Release Management Change Management Distributed Systems
7+ yrs exp

Principal System Engineering

AT&T

Alpharetta, GA +4 9 days ago $155,400$261,100
Kafka IXBUS Microservices CI/CD Observability Monitoring Logging Cybersecurity Disaster Recovery Capacity Planning Cost Optimization Automation Event-Driven Architecture
7+ yrs exp

Principal System Engineering

AT&T

Atlanta, GA +5 9 days ago $155,400$261,100
Kafka IXBUS Microservices CI/CD Observability Disaster Recovery Capacity Planning Cost Optimization Automation
7+ yrs exp

Principal System Engineering

AT&T

Alpharetta, GA +4 9 days ago $155,400$261,100
Kafka IXBUS Microservices CI/CD Observability Monitoring Logging Cybersecurity Disaster Recovery Capacity Planning Cost Optimization Automation Event-Driven Architecture
7+ yrs exp