Senior Product Manager –AI Inference Software

AMD - Santa Clara, CA

Hiring: Senior Product Manager –AI Inference Software Company: AMD Location: Santa Clara, CA Job Posted Time: 2026-09-09 18:02:28 Target Skills & Keywords : GitHub, HBase About the job Required Skills: •This role owns the framework-layer inference product strategy for ROCm, translating customer, ecosystem, and engineering signals into roadmap decisions for production AI inference on AMD Instinct hardware. As inference becomes the defining workload for production AI, you will help shape how AMD’s software ecosystem enables efficient, reliable, and competitive large-scale model deployment. You will work with engineering, strategic AI customers, ecosystem partners, and the open-source inference community to advance inference at scale. •You are a technically deep product leader with strong expertise in AI inference infrastructure and open-source software. You can reason across inference engines, serving, orchestration, memory management, and performance while understanding how these systems interact with the GPU software and hardware beneath them. You navigate complex organizations, build alignment without formal authority, and drive important work to completion. You are comfortable operating at the intersection of open-source communities and enterprise-scale customers, whether digging into a GitHub issue thread or presenting roadmap tradeoffs to a VP of Engineering at a hyperscaler. •Key Responsbilities •Product Strategy & Roadmap •Own the product strategy and roadmap for ROCm’s inference frameworks software capabilities. •Define how AMD’s framework-layer software stack for inference enables production workloads from single-GPU through rack-scale deployments, with a focus on performance, developer experience, portability, and competitive differentiation. •Identify emerging model architectures, serving technologies, and infrastructure shifts that require changes in the inference frameworks layer and translate them into product priorities. •Balance customer needs, ecosystem direction, technical opportunities, and engineering investment to determine where AMD should differentiate in the software stack. •Inference Software & Engineering Partnership •Partner with engineering teams across inference engines, serving and orchestration, memory systems, GPU libraries, runtimes, kernels, and communications, with particular emphasis on the framework and orchestration layers that enable large-scale inference above the underlying GPU software stack. •Translate customer and ecosystem needs into prioritized engineering requirements and drive execution across organizational boundaries. •Work with engineering Compensation: •$205,680 •$308,520 Interested candidates, please apply directly through the job posting on company's career page or try via AI auto apply on this platform. Don't miss this opportunity to join a forward-thinking team!