Site Reliability Engineer, Infrastructure Platforms

GitLab

Confirmed live yesterday High trust
Remote

Quick summary

Work type
Remote
Location
United Kingdom
Posted
3 days ago
Freshness
Confirmed live yesterday

Market check

Salary context

How this pay compares to similar roles

Similar $179k
$133k most similar roles pay here $229k

This listing doesn't post a salary. Most similar roles pay $144,350–$214,000.

Based on 240 similar postings.

Employer

About GitLab

GitLab is an all-remote software company that develops an AI-powered DevSecOps platform combining source code management, CI/CD, security scanning, and project planning in a single application.

GitLab currently has 79 open roles on FindRole.

Listed pay typically runs $137,400–$213,600 across 61 roles with salary data.

Most-posted roles

View all roles at GitLab

At a glance

TL;DR · Site Reliability Engineer, Infrastructure Platforms

Site Reliability Engineer, Infrastructure Platforms — UK (Intermediate to Senior Staff) joins the Infrastructure Platforms team to ensure user-facing services and production systems remain reliable, scalable, and efficient. The role involves building automation and tooling to reduce toil through infrastructure-as-code workflows, managing Kubernetes deployments, and maintaining CI/CD pipelines via GitOps. You will contribute to the observability stack using metrics, logs, and SLOs while participating in on-call rotations and incident response. Key technical requirements include proficiency in Go or Ruby, experience with Terraform modules, and expertise in Kubernetes and its ecosystem. Candidates must have hands-on experience with major cloud providers like GCP or AWS. The role focuses on solving complex reliability challenges for the GitLab.com platform by transforming manual tasks into repeatable processes through software engineering principles and automated systems to maintain high availability at scale.

What does a Site Reliability Engineer earn?

Median $186200 from 131 postings across 36 companies.

See salary data

What you'll do

  • Maintain the reliability, scalability, and efficiency of user-facing production systems.
  • Build automation and tooling to replace manual work with infrastructure-as-code workflows.
  • Operate and troubleshoot production systems on Kubernetes, including deployments and scaling.
  • Write and maintain infrastructure as code using CI/CD and GitOps methodologies.
  • Participate in on-call rotations, triage alerts, and improve existing runbooks.
  • Contribute to the observability stack using metrics, logs, and SLOs to detect issues early.
  • Lead incident response and conduct post-incident reviews to improve automation and processes.
  • Document architecture decisions and technical findings to create repeatable practices.

What we're looking for

  • Experience keeping production systems reliable by combining an operations mindset with software engineering practices.
  • Experience building net-new infrastructure tooling and automation, such as Terraform modules or Kubernetes operators.
  • Ability to read, debug, and reason about code in Go or Ruby.
  • Experience with infrastructure as code and the Kubernetes ecosystem.
  • Hands-on experience with at least one major cloud provider, specifically GCP or AWS.
  • Familiarity with observability practices including metrics, logging, alerting, and SLOs/SLIs.
  • Comfort participating in on-call rotations and incident response under pressure.
  • Strong written communication skills for operating in an asynchronous, distributed environment.

More like this

Similar roles

Site Reliability Engineer

Balyasny Asset Management

Warsaw, Poland 78 days ago
Prometheus Grafana Loki Tempo OTEL Kubernetes Docker AWS Python Bash Go CI/CD DevOps SRE Agile
5+ yrs exp

Site Reliability Engineer

Booz Allen Hamilton

McLean, VA 9 days ago $86,800$198,000
AWS Kubernetes Terraform Ansible CI/CD Python Bash PowerShell Docker Prometheus Grafana Loki Elasticsearch Kibana GitLab GitHub CloudFormation OpenShift Jenkins REST JSON YAML XML Agile
6+ yrs exp

Senior Site Reliability Engineer, Infra Ops

Circle

Remote (San Francisco, CA) 28 days ago $152,500$205,000
Kubernetes Terraform Go Python JavaScript TypeScript CI/CD Infrastructure as Code Distributed Systems Cloud Infrastructure Observability GitOps SRE DevOps Networking Security incident-management Capacity Planning
5+ yrs exp Remote