Manager, Site Reliability Engineering

Mastercard

Confirmed live yesterday High trust

Quick summary

Work type
On-site
Location
O Fallon, IL
Salary
$122,000–$207,000 / yr
Posted
16 days ago
Freshness
Confirmed live yesterday

Market check

Salary context

Below market

How this pay compares to similar roles

Similar $184k
This role $164k
$109k most similar roles pay here $243k

This role pays less than 68% of similar roles. Most pay $152,150–$215,000 — the shaded band above. At the midpoint, this role pays about $164k versus about $184k for comparable roles.

Based on 238 similar postings.

Employer

About Mastercard

Mastercard is a global technology company in the payments industry, processing transactions between financial institutions and merchants using its extensive network of credit, debit, and prepaid card products. Industry: Payments Technology & Financial Services

Mastercard currently has 116 open roles on FindRole.

Listed pay typically runs $122,000–$207,000 across 104 roles with salary data.

Most-posted roles

View all roles at Mastercard

At a glance

TL;DR · Manager, Site Reliability Engineering

Manager, Site Reliability Engineering joins the Business Operations team as a production readiness steward for Mastercard products. This role focuses on ensuring platform stability through monitoring, automated change implementation, and operational design to create fault-tolerant, scalable systems. The individual will manage daily operations with a focus on triage, root cause analysis, and blameless post-mortems while managing risk and compliance across all environments. Key responsibilities include serving as an Operational Readiness Architect, performing capacity planning, and developing monitoring strategies to achieve zero downtime during deployment. Required skills include experience with Splunk and Dynatrace, understanding of network concepts, stack trace analysis, and high availability planning. The candidate should possess a degree in Computer Science or a related field, coding and scripting proficiency, and an ability to troubleshoot large-scale distributed systems while managing data integrity and information governance.

What you'll do

  • Serve as the primary contact responsible for overall application health, performance, and capacity.
  • Design monitoring and alerting strategies to ensure zero downtime during deployments.
  • Develop automation and tooling to reduce manual intervention and operational toil.
  • Perform triage and conduct blameless post-mortems to identify root causes of production issues.
  • Manage risk by ensuring compliance and security across all environments.
  • Consult on system design, capacity planning, and launch reviews for new services.
  • Analyze ITSM activities to provide feedback to development teams regarding operational gaps.
  • Enforce information governance policies to maintain data integrity and regulatory compliance.

What we're looking for

  • Bachelor of Science degree in Computer Science or a related technical field involving coding, or equivalent practical experience.
  • Experience with coding or scripting.
  • Basic to intermediate understanding of algorithms, data structures, scripting, pipeline management, and software design.
  • Experience with monitoring tools such as Splunk and Dynatrace.
  • Understanding of network concepts, stack trace analysis, high availability, and business continuity planning.
  • Ability to troubleshoot large-scale distributed systems.
  • Strong communication skills and a systematic approach to problem-solving.
  • Ability to collaborate with cross-functional teams in a geographically distributed environment.

More like this

Similar roles

Senior Site Reliability Engineer

Mastercard

O Fallon, IL 29 days ago $96,000$163,000
Site Reliability Engineering Python Go Bash Linux Unix AWS Azure GCP CI/CD Containerization Orchestration Observability Monitoring Troubleshooting Capacity Planning Network Administration IT Service Management

Site Reliability Engineer II

Mastercard

O Fallon, IL 29 days ago $76,000$127,000
Python Go Bash Linux Unix AWS Azure GCP CI/CD Containerization Orchestration Observability Monitoring Troubleshooting Capacity Planning IT Service Management Network Administration

Lead Site Reliability Engineer

Mastercard

O Fallon, IL 129 days ago $122,000$207,000
Java Spring Framework Python Go Site Reliability Engineering DevOps CI/CD Distributed Systems Automation Configuration Management Observability Root Cause Analysis ITSM Capacity Planning Monitoring

Senior Site Reliability Engineer

Mastercard

O Fallon, IL 113 days ago $96,000$163,000
Unix Shell Scripting SQL Python Apache NiFi Splunk Dynatrace Jenkins Git CI/CD C C++ Java Go Perl Ruby Maven Artifactory Chef BitBucket

Manager, Site Reliability Engineering

Okta Inc

San Francisco, CA 43 days ago $204,000$306,000
AWS Kubernetes Terraform CI/CD Grafana Splunk APM DevOps SaaS Cloud-native Architecture Containerization Edge Networking SDLC
3+ yrs exp Hybrid

Manager, Site Reliability Engineering

Oracle

Reston, VA +1 50 days ago $102,000$234,600
Site Reliability Engineering Incident Response Automation Monitoring Provisioning Decommissioning Scalability Data Collection
6+ yrs exp