Lead Site Reliability Engineer, Data

Comcast

Confirmed live yesterday Trusted

Quick summary

Work type
On-site
Location
Reston, VA
Salary
$152,020–$228,030 / yr
Posted
79 days ago
Freshness
Confirmed live yesterday

Market check

Salary context

Competitive pay

How this pay compares to similar roles

Similar $181k
This role $190k
$134k most similar roles pay here $238k

This role pays more than 52% of similar roles. Most pay $147,158–$214,625 — the shaded band above. At the midpoint, this role pays about $190k versus about $181k for comparable roles.

Based on 240 similar postings.

Employer

About Comcast

Comcast is an American telecommunications and media conglomerate, providing cable TV, internet, and phone services under the Xfinity brand, and owning NBCUniversal.

Comcast currently has 84 open roles on FindRole.

Listed pay typically runs $137,369–$206,053 across 58 roles with salary data.

Most-posted roles

View all roles at Comcast

At a glance

TL;DR · Lead Site Reliability Engineer, Data

Lead Site Reliability Engineer, Data- FreeWheel joins the Global Operation team to ensure the reliability, scalability, and performance of data systems. This role involves managing data infrastructure, optimizing system reliability, automating daily operations, and resolving technical issues affecting data pipelines and backend platforms. The engineer will design monitoring and alerting systems, develop automation tools for deployment and disaster recovery, perform capacity planning, and manage security compliance. Key responsibilities include troubleshooting platform failures and collaborating with data science and product teams to support data product designs. Required skills include experience with cloud platforms like AWS, GCP, or Azure; modern architectures including Kafka, Hadoop, Spark, Cassandra, HDFS, and S3; databases such as NoSQL, MySQL, and PostgreSQL; tools like Ansible, Terraform, Kubernetes, Docker, Prometheus, Grafana, and the ELK Stack. Programming proficiency in Python, Go, Java, or Scala is required.

What you'll do

  • Design and implement monitoring and alerting systems to ensure data platform stability and performance.
  • Develop and maintain automation scripts for deployment, backup, recovery, and disaster recovery of data systems.
  • Analyze and optimize the performance of data storage, query execution, and large-scale data flows.
  • Respond to and troubleshoot failures in data pipelines and backend platforms to ensure high availability.
  • Forecast capacity requirements and scale infrastructure to accommodate growing data volumes.
  • Document architecture, configurations, and operational procedures for internal knowledge sharing.
  • Ensure all data platforms meet security standards and compliance requirements to prevent unauthorized access.

What we're looking for

  • At least 10 years of experience as an SRE, DevOps, or Data Operations Engineer.
  • Bachelor's degree or higher in Computer Science, Software Engineering, or a related field.
  • Experience with cloud platforms including AWS, GCP, and Azure.
  • Familiarity with modern data architectures such as Kafka, Hadoop, Spark, Cassandra, HDFS, and AWS S3.
  • Extensive experience in database management including NoSQL, MySQL, and PostgreSQL.
  • Proficiency with Ansible, Terraform, Kubernetes, and Docker.
  • Programming skills in Python, Go, Java, or Scala.
  • Experience with monitoring tools such as Prometheus, Grafana, or the ELK Stack.

More like this

Similar roles

Senior Site Reliability Engineer, Data

Comcast

Reston, VA 79 days ago $128,830$193,245
AWS GCP Azure Kafka Hadoop Spark Cassandra HDFS S3 NoSQL MySQL PostgreSQL Ansible Terraform Kubernetes Docker CI/CD Python Go Java Scala Prometheus Grafana ELK Stack Snowflake Aerospike
8+ yrs exp

Senior Site Reliability Engineer, Data

Comcast

Chicago, IL 79 days ago $117,627$176,441
AWS Kubernetes Terraform Python Go Java Scala Docker Prometheus Grafana ELK Stack Ansible MySQL PostgreSQL NoSQL CI/CD Microservices ETL Pipelines
8+ yrs exp

Site Reliability Engineer, Lead

Booz Allen Hamilton

Chantilly, VA 50 days ago $99,000$225,000
Prometheus Grafana ELK Stack Linux AWS Python Terraform Terragrunt Kubernetes OpenTelemetry AWS CloudWatch AWS EKS Rancher Jenkins Git Docker Nessus JIRA Confluence SRE
8+ yrs exp

Senior Lead Site Reliability Engineer

JPMorgan Chase

Palo Alto, CA 59 days ago
Site Reliability Engineering Java Go Python Terraform Kubernetes Docker CI/CD GitOps Grafana Prometheus Dynatrace Datadog Splunk Kafka RabbitMQ SQS Neo4j Pinecone Weaviate Chroma LangChain LangGraph AutoGen CrewAI GitHub Copilot Fluentd Logstash Vector RESTful APIs RAG TensorFlow PyTorch scikit-learn Hadoop Spark Flink MongoDB Cassandra DynamoDB InfluxDB TimescaleDB AWS Azure GCP
5+ yrs exp

Senior Lead Site Reliability Engineer

JPMorgan Chase

Plano, TX 45 days ago
Site Reliability Engineering Observability Monitoring Telemetry Service Level Objectives Alerting AI SDLC Automation
5+ yrs exp