AI Instinct System Management Architect

Amd

Confirmed live yesterday High trust
Hybrid

Quick summary

Work type
Hybrid
Location
Santa Clara, CA
Salary
$237,200–$355,800 / yr
Posted
23 days ago
Freshness
Confirmed live yesterday
Closes
Aug 19, 2027

Market check

Salary context

Above market

How this pay compares to similar roles

Similar $200k
This role $296k
$132k most similar roles pay here $380k

This role pays more than 95% of similar roles. Most pay $162,000–$237,925 — the shaded band above. At the midpoint, this role pays about $296k versus about $200k for comparable roles.

Based on 239 similar postings.

Employer

About Amd

AMD (Advanced Micro Devices) is a semiconductor company that develops high-performance processors, graphics cards, and adaptive computing solutions for gaming, data centers, and embedded markets. Industry: Semiconductors

Amd currently has 367 open roles on FindRole.

Listed pay typically runs $166,400–$249,600 across 367 roles with salary data.

Most-posted roles

View all roles at Amd

At a glance

TL;DR · AI Instinct System Management Architect

The AI Instinct System Management Architect will define and drive the architecture for system management and observability across AI datacenter platforms. This role involves designing unified architectures for rack-scale and pod-scale management, spanning firmware, operating systems, rack controllers, and orchestration layers to create manageable, composable, and secure infrastructure. The architect will develop telemetry frameworks, manage lifecycle workflows like provisioning and firmware upgrades, and produce reference designs for customer integration. Key technical requirements include expertise in BMC firmware stacks, Redfish and DMTF standards, and experience with PCIe, CXL, NVMe interconnects, and cluster schedulers like Kubernetes or Slurm. Preferred skills include familiarity with Prometheus, Loki, ELK, and OpenTelemetry, along with proficiency in C/C++, Python, or Go. The role addresses the technical challenge of managing complex infrastructure for large-scale AI workloads.

What you'll do

  • Define and drive the architecture for system management and observability across AMD's AI datacenter platforms.
  • Develop unified architectures for rack-scale and pod-scale management from firmware to orchestration layers.
  • Ensure compatibility with industry standards like Redfish and DMTF profiles for monitoring and analytics.
  • Deliver manageability solutions including BMC/BSP, rack/pod controllers, and APIs for large-scale infrastructure.
  • Create telemetry frameworks and standard-based interfaces for compute, storage, networking, and accelerators.
  • Manage lifecycle workflows such as discovery, provisioning, firmware upgrades, and decommissioning across racks and pods.
  • Collaborate with customers to align designs with DCIM/ITSM environments and lead proof-of-concept validations.
  • Produce architectural collateral including reference designs, integration guides, and telemetry baselines for customer readiness.

What we're looking for

  • Expert background in systems or platform software architecture with a focus on system management and server manageability.
  • Deep expertise in BMC firmware stacks, telemetry, inventory, alerting, and management protocols.
  • Strong knowledge of DMTF standards (MCTP, PLDM, SPDM, Redfish), platform security, and management networking.
  • Experience with PCIe, CXL, NVMe interconnects and cluster schedulers like Kubernetes or Slurm.
  • Proven ability to combine technical leadership with customer engagement for scalable AI datacenter deployments.
  • Experience with GPU/accelerator platforms (preferred).
  • Familiarity with telemetry stacks such as Prometheus, Loki, ELK, or OpenTelemetry (preferred).
  • Knowledge of datacenter infrastructure components and strong programming skills in C/C++, Python, or Go (preferred).

More like this

Similar roles

Director, AI Instinct Stack Architecture

Amd

Austin, TX 21 days ago $216,640$324,960
GPU HPC System Architecture Firmware System Software Hardware/Software Co-design RAS Telemetry Platform Management Security Diagnostics SoC
Hybrid

AI Systems Security Architect

Amd

San Diego, CA 7 days ago $174,400$261,600
C C++ Assembly Root of Trust Secure Boot Measured Boot Confidential Computing Trusted Execution Environments TPM TCG DICE SPDM PCIe CXL SR-IOV JTAG SBOM NIST OpenBMC coreboot
Hybrid

AI Data Center System Architect

Amd

Santa Clara, CA +2 2 days ago $229,600$344,400
System Architecture AI Infrastructure High-Performance Computing Silicon Thermal Architecture Power System Design High-Speed Electrical Interfaces Enterprise AI

AI Platform Architect

Amd

Austin, TX +2 25 days ago $229,600$344,400
AI Infrastructure Rack-scale Systems Hardware/Software Co-design Firmware PCIe Ethernet NVMe Redfish IPMI SMBus System Integration Platform Security

AI Solution Architect

Booz Allen Hamilton

Atlanta, GA 8 days ago $99,000$225,000
RAG Prompt Design Agentic Workflows LangChain Semantic Kernel AutoGen Model Context Protocol Azure AI Foundry GitHub Actions CI/CD NodeJS .NET SQL Infrastructure-as-Code Containerization GitHub Copilot Claude Code
5+ yrs exp

AI Solution Architect

Booz Allen Hamilton

Hanscom AFB, MA +1 2 days ago $112,900$257,000
AI LLM Palantir Foundry AWS Kubernetes DevSecOps CI/CD Containerization Data Pipelines Ontologies Automated Testing Analytics Monitoring Frameworks Agentic Workflows
10+ yrs exp