Server Platform Debug Program Lead

Advanced Micro Devices, Inc•Austin, TX
•Hybrid

About The Position

As the Debug Program Engineer, you will own the operating cadence of the Server Platform Debug Steering Forum — the single forum where platform, silicon, firmware, and validation teams converge to drive open debug issues to closure. You will be accountable for the accuracy and integrity of debug execution tracking across all active server programs, for holding technical owners accountable to closure plans, and for securing the hardware and technical expertise required to unblock investigations. This is a hands-on technical program role: you will be expected to understand the failure signatures you are tracking, not simply transcribe them. The right candidate is a technically knowledgeable and engineer with a working foundation in server platform or SOC debug — someone who can read a triage update, assess whether it is credible, and challenge a stalled root-cause path. They will have the communication skills and leadership ability to coordinate deliverables from many diverse technical leads. They are comfortable operating between engineering and program management, driving urgency without formal authority, and escalating with evidence.

Requirements

  • Direct exposure to server platform, SOC, or board-level debug (boot failures, signal integrity, memory training, thermal, power sequencing, or firmware-related issues)
  • Familiarity with debug tooling and triage flows, and the ability to interpret logs, failure signatures, and root-cause analyses
  • Excellent verbal and written communication skills, with the ability to summarize deeply technical status for executive audiences
  • Flexible working schedule to manage a global debug execution team (Asia and North America emphasis)

Nice To Haves

  • Experience running a recurring technical review or triage forum is a strong plus

Responsibilities

  • Lead the Server Platform Debug Steering Forum end to end — agenda, prioritization, technical review of open issues, decisions, and follow-through between sessions
  • Maintain an accurate, single source of truth for debug execution: issue inventory, severity and priority, technical owner, current hypothesis, next debug step, and expected closure date
  • Ensure technical owners are actively driving to root cause and closure; identify stalled or circular investigations early and escalate with a clear ask
  • Coordinate the resources required to sustain debug momentum — system/board availability, instrumentation, lab and bench time, silicon samples, and access to subject-matter experts across firmware, SI/PI, thermal, and validation
  • Drive cross-team dependencies where an issue spans platform, silicon, firmware, or supplier scope, and broker the handoffs needed to keep the investigation moving
  • Distinguish genuine technical blockers from resourcing or prioritization gaps, and route each to the right decision-maker
  • Publish recurring debug status to program and engineering leadership — burn-down, aging, top blockers, and launch-risk issues — via dashboards and written summaries
  • Capture and propagate debug learnings, known-issue patterns, and triage best practices across programs
  • Operate with urgency during escalations and critical launch-blocking debug efforts

Benefits

  • AMD benefits at a glance
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service