This role owns the framework-layer inference product strategy for ROCm, translating customer, ecosystem, and engineering signals into roadmap decisions for production AI inference on AMD Instinct hardware. As inference becomes the defining workload for production AI, you will help shape how AMD’s software ecosystem enables efficient, reliable, and competitive large-scale model deployment. You will work with engineering, strategic AI customers, ecosystem partners, and the open-source inference community to advance inference at scale. You are a technically deep product leader with strong expertise in AI inference infrastructure and open-source software. You can reason across inference engines, serving, orchestration, memory management, and performance while understanding how these systems interact with the GPU software and hardware beneath them. You navigate complex organizations, build alignment without formal authority, and drive important work to completion. You are comfortable operating at the intersection of open-source communities and enterprise-scale customers, whether digging into a GitHub issue thread or presenting roadmap tradeoffs to a VP of Engineering at a hyperscaler.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Senior