Senior Lead Software Engineer, AI Platforms

JPMorgan Chase

Confirmed live today High trust

Quick summary

Work type
On-site
Location
Seattle, WA
Posted
4 days ago
Freshness
Confirmed live today

Market check

Salary context

How this pay compares to similar roles

Similar $199k
$140k most similar roles pay here $262k

This listing doesn't post a salary. Most similar roles pay $151,991–$246,150.

Based on 240 similar postings.

Employer

About JPMorgan Chase

JPMorgan Chase & Co. is a global financial services firm and one of the largest banks in the world, offering investment banking, commercial banking, asset management, and consumer financial services.

JPMorgan Chase currently has 1138 open roles on FindRole.

Most-posted roles

View all roles at JPMorgan Chase

At a glance

TL;DR · Senior Lead Software Engineer, AI Platforms

As a Senior Lead Software Engineer, AI Platforms within the Infrastructure Platforms team, you will design, build, and operate secure, scalable infrastructure platforms specifically optimized for multi-GPU and multi-node machine learning training workloads. You will collaborate with AI/ML engineering teams to translate complex compute, storage, networking, and GPU requirements into production-ready capabilities while driving automation and developer productivity through the software delivery lifecycle. Your daily work involves developing high-quality code, managing CI/CD pipelines, and implementing infrastructure-as-code solutions. The role requires expertise in Kubernetes, Docker, and programming languages like Python, Go, Java, or C#. You will also manage GPU infrastructure resources and oversee AI-assisted engineering practices. Technical focus areas include distributed systems, high-performance computing, MLOps tools like MLflow, and specialized NVIDIA GPU software ecosystems to solve complex scalability and reliability challenges in enterprise environments.

What you'll do

  • Architect and operate secure, scalable cloud infrastructure platforms optimized for multi-GPU and multi-node AI/ML training workloads.
  • Translate complex compute, storage, networking, and GPU requirements into production-ready infrastructure capabilities.
  • Develop, test, and deliver high-quality production code while performing peer code reviews and debugging.
  • Build and maintain CI/CD pipelines and infrastructure-as-code solutions to automate ML platform deployment and operations.
  • Monitor and optimize cloud resources for performance, reliability, utilization, scalability, and cost efficiency.
  • Lead technical design decisions regarding product architecture, infrastructure strategy, and operational effectiveness.
  • Drive the adoption of approved AI-assisted engineering practices to improve code quality and delivery speed.
  • Provide technical leadership and guidance to engineers and partners to ensure alignment with security and engineering standards.

What we're looking for

  • Formal training or certification in software engineering concepts and 5+ years of applied experience.
  • Hands-on experience building scalable infrastructure for machine learning training and inference workloads.
  • System-level understanding of GPU infrastructure, accelerators, high-speed interconnects, and distributed compute technologies.
  • Strong experience with Kubernetes, Docker, containerization, and production troubleshooting.
  • Proficiency in at least one modern programming language such as Python, Go, Java, or C#.
  • Deep understanding of cloud architecture including microservices, storage, networking, security, routing, and switching.
  • Demonstrated experience leading the use of enterprise-authorized AI-assisted software development tools for engineering productivity.
  • Experience with NVIDIA GPU infrastructure, MLOps platforms, high-performance computing, and distributed systems (preferred).

More like this

Similar roles

Senior Lead Software Engineer, AI Platform Engineer

JPMorgan Chase

Seattle, WA 4 days ago
Kubernetes Docker Python Go Java C# Infrastructure as Code CI/CD Cloud Computing Microservices Machine Learning ML Ops Prometheus Grafana MLflow vLLM Ray.io Slurm SQL NoSQL Linux
5+ yrs exp

Senior Lead Software Engineer, AI/ML Platform

JPMorgan Chase

Wilmington, DE 31 days ago
Python Java Go AWS Kubernetes Terraform Docker CI/CD MLOps Kubeflow MLflow PyTorch TensorFlow Hugging Face scikit-learn SQL NoSQL Linux Microservices Distributed Systems
5+ yrs exp

Lead Software Engineer

JPMorgan Chase

Jersey City, NJ 20 days ago
Python Java Go AWS Kubernetes Terraform Docker CI/CD MLOps PyTorch TensorFlow Hugging Face scikit-learn SQL NoSQL Linux Microservices infrastructure-as-code Kubeflow MLflow
5+ yrs exp

Senior Software Engineer, AI Platform

Anduril Industries

Seattle, WA 63 days ago $191,000$253,000
CI/CD Python C++ Rust Terraform Ansible Docker Kubernetes GitLab CI Jenkins GitHub Actions Crossplane Puppet Chef Bash Prometheus Grafana ELK Stack Splunk GitOps