Lead Product Engineer, AIOps for Observability

Allstate

Confirmed live 2 days ago High trust
Remote

Quick summary

Work type
Remote
Location
IL
Salary
$100,000–$170,500 / yr
Posted
23 days ago
Freshness
Confirmed live 2 days ago

Market check

Salary context

Below market

How this pay compares to similar roles

Similar $185k
This role $135k
$86k most similar roles pay here $229k

This role pays less than 82% of similar roles. Most pay $154,656–$215,550 — the shaded band above. At the midpoint, this role pays about $135k versus about $185k for comparable roles.

Based on 240 similar postings.

Employer

About Allstate

The Allstate Corporation is one of the largest publicly held personal lines insurers in the US, widely recognized for its "You're In Good Hands With Allstate®" slogan.

Allstate currently has 36 open roles on FindRole.

Listed pay typically runs $100,000–$170,500 across 36 roles with salary data.

Most-posted roles

View all roles at Allstate

At a glance

TL;DR · Lead Product Engineer, AIOps for Observability

Lead Product Engineer – (AIOps for Observability) serves as a technical leader within the AI-powered Observability platform team. The role focuses on designing and developing next-generation observability solutions powered by agentic AI to enable proactive detection, diagnosis, and automated remediation across hybrid and multi-cloud environments. You will build scalable platforms, developer-centric tooling, and AIOps frameworks that integrate LLMs and orchestration tools to reduce noise and improve incident response. Key responsibilities include creating self-healing systems, managing logs, metrics, traces, and events, and enhancing platforms like Datadog, Dynatrace, New Relic, AppDynamics, and OTEL. Required technical skills include Java/Spring Boot, Python, Node.js, React, Kubernetes, and cloud-native architectures. The role addresses the challenge of improving reliability and operational efficiency by automating manual workflows and providing real-time health analytics for complex enterprise systems.

What you'll do

  • Design and build agentic AI solutions for autonomous anomaly detection, root cause analysis, and automated remediation.
  • Integrate LLMs and AI orchestration frameworks to create self-service diagnostics and intelligent alerting systems.
  • Architect end-to-end observability solutions across logs, metrics, traces, and events in hybrid and multi-cloud environments.
  • Enhance existing observability platforms like Datadog, Dynatrace, and New Relic with custom integrations and AI capabilities.
  • Develop full-stack tools and APIs using Java, Python, or Node.js to simplify instrumentation and monitoring for developers.
  • Build self-healing systems that automatically detect and resolve infrastructure issues to improve operational efficiency.
  • Implement infrastructure-as-code and automated deployment pipelines to ensure system reliability and scalability.
  • Define enterprise observability architecture while mentoring teams on AI and engineering best practices.

What we're looking for

  • 5+ years of experience building software, platforms, or automation solutions with a focus on observability and distributed systems.
  • 4+ years of hands-on experience with observability platforms such as Datadog, Dynatrace, or OTEL.
  • 3+ years of software development experience using Java (Spring Boot), Python, Node.js, or React.
  • 2+ years of hands-on experience building agentic AI systems, including autonomous agents and AI workflows.
  • Proven ability to design and implement AI-driven automation and AIOps solutions.
  • Experience integrating LLMs, orchestration frameworks, or AI pipelines into production systems.
  • Strong experience with Kubernetes, cloud-native architectures, and hybrid environments (on-prem and cloud).
  • Expertise in API development, event-driven systems, microservices, and CI/CD pipelines.

More like this

Similar roles

Senior Observability Automation Engineer

Q2

Austin, TX 11 days ago
C# Go Python Bash PowerShell Perl RESTful API CI/CD Infrastructure as Code Observability Grafana Splunk Cloud Infrastructure DevOps Linux Windows Event-driven Automation AI Platform
8+ yrs exp

Lead Software Engineer, Python, Observability

JPMorgan Chase

Houston, TX 49 days ago
Python SQL LLM RAG Agentic Systems CI/CD OTEL Grafana Splunk Dynatrace Docker Kubernetes Airflow Agile Monitoring Observability Version Control
8+ yrs exp

Lead Engineer, Observability Platform

CVS Health

Remote (RI) +4 25 days ago $106,605$284,280
OpenTelemetry Go Python Java Spring Boot Kubernetes Docker Argo CD AWS GCP Azure Terraform CloudFormation Helm Kustomize Grafana Loki Tempo Mimir PostgreSQL MySQL Kafka Pulsar ClickHouse Cassandra Istio Envoy CI/CD
7+ yrs exp Remote

Senior Software Engineer, AIOps and Observability

Nvidia

Santa Clara, CA 42 days ago $200,000$322,000
AIOps Observability Prometheus Victoria Metrics Vector Loki Grafana Alert Manager Clickhouse OpenTelemetry BigPanda PagerDuty Datadog Kubernetes Nomad Docker Microservices NATS Kafka Go Python Java C# Machine Learning Generative AI LLMs
10+ yrs exp

Principal Observability Automation Engineer

Q2

Austin, TX 14 days ago
C# Go Python Bash PowerShell AWS Azure GCP VMware Docker Terraform CI/CD SQL NoSQL MS SQL Server Grafana Splunk NGINX IIS Linux Windows Server RESTful APIs Microservices
10+ yrs exp

Senior Research & Development Engineer

Medtronic

Remote (San Diego, CA) 3 days ago $105,600$158,400
Creo SolidWorks Systems Engineering Windchill Agile Additive Manufacturing Rapid Prototyping Verification and Validation New Product Introduction Quality System Regulation
4+ yrs exp Remote