Senior Site Reliability Engineer, Capacity, Platform Infrastructure

Elastic

Confirmed live yesterday High trust

Quick summary

Work type
On-site
Location
Posted
17 days ago
Freshness
Confirmed live yesterday

Market check

Salary context

How this pay compares to similar roles

Similar $182k
$119k most similar roles pay here $225k

This listing doesn't post a salary. Most similar roles pay $151,000–$212,375.

Based on 240 similar postings.

Employer

About Elastic

Elastic is the company behind Elasticsearch and the Elastic Stack (Elasticsearch, Kibana, Beats, and Logstash). It sells search, observability, and security products built on that search engine, primarily through its Elastic Cloud managed service.

Elastic currently has 240 open roles on FindRole.

Listed pay typically runs $133,100–$210,600 across 91 roles with salary data.

Most-posted roles

View all roles at Elastic

At a glance

TL;DR · Senior Site Reliability Engineer, Capacity, Platform Infrastructure

Senior Site Reliability Engineer (Capacity) - Platform Infrastructure joins the platform engineering team to manage and optimize compute resources for Elastic Cloud Hosted and Serverless workloads. This role focuses on ensuring seamless scaling by assessing future capacity requirements, developing accurate prediction models, and implementing strategies to optimize resource usage across cloud environments. The engineer will analyze metrics to guide allocation decisions, build reporting tools for visibility, and operate an autoscaling framework across over 60 regions. Key responsibilities include troubleshooting infrastructure performance and managing capacity reservations across the three major cloud service providers. Candidates should possess a strong software and platform engineering background with expertise in performance monitoring, incident investigation, and navigating complex cloud scaling challenges to solve real-world resource allocation problems within a multi-region cloud infrastructure environment.

What does a Site Reliability Engineer earn?

Median $186200 from 131 postings across 36 companies.

See salary data

What you'll do

  • Assess current and future capacity requirements to ensure seamless scaling of cloud resources.
  • Develop and maintain accurate capacity models to predict resource needs based on business objectives.
  • Implement strategies to optimize compute resource usage across various cloud environments.
  • Analyze capacity metrics and trends to guide informed resource allocation decisions.
  • Create reporting tools to provide visibility into infrastructure performance and capacity.
  • Operate an autoscaling framework to accommodate diverse customer workloads across multiple regions.
  • Manage compute capacity reservations and scaling issues across the three major cloud service providers.

What we're looking for

  • 5+ years of experience with cloud infrastructure and capacity management.
  • Knowledge of performance monitoring and optimization techniques.
  • Understanding of cloud scaling challenges and solutions.
  • Proficiency with incident investigation and troubleshooting processes.
  • Experience with compute auto-scaling processes and capacity reservations across the three major CSPs.
  • Solid software and platform engineering background.
  • Experience working with the three major cloud service providers to navigate compute capacity scaling issues.
  • Export controls.

More like this

Similar roles

Senior Site Reliability Engineer

Autodesk

Remote (ID) +1 28 days ago $117,000$209,330
Site Reliability Engineering AWS Kubernetes Python Go Java Infrastructure as Code CI/CD CloudWatch Splunk Datadog Dynatrace Bash PowerShell FedRAMP Distributed Systems Load Balancing DNS
7+ yrs exp Remote

Senior Site Reliability Engineer

Autodesk

San Francisco, CA 28 days ago $117,000$209,330
SRE Python Go Java Bash PowerShell AWS Kubernetes Infrastructure as Code CI/CD CloudWatch Splunk Datadog Dynatrace FedRAMP Distributed Systems Load Balancing DNS
7+ yrs exp

Senior Site Reliability Engineer

Salesforce

Remote (San Francisco, CA) 58 days ago $148,500$223,900
SRE Python Go Docker Kubernetes CI/CD Prometheus Grafana ELK Splunk Datadog Temporal Airflow Argo Workflows AWS GCP Linux Unix LLM Prompt Engineering
5+ yrs exp Remote