About The Position

Cerebras Systems is seeking an experienced Senior Technical Program Manager to establish an accountable operating layer across feature delivery, release integration testing, Core Infrastructure, branch stability, and release readiness within the AI Inference Core. This role involves creating a clear, measurable, and consistently followed operating mechanism for complex technical execution. The successful candidate will ensure cross-team initiatives have defined ownership, entry/exit criteria, visible health and SLA tracking, timely escalation, and dependable follow-through. This is a strategic role requiring technical depth to understand complex AI systems, challenge plans, identify dependencies, improve decision quality, and build trusted operating mechanisms for engineering teams, rather than a simple project tracking or meeting coordination function.

Requirements

  • Significant experience leading complex technical programs across multiple engineering teams, ideally in infrastructure, distributed systems, platforms, release engineering, or AI systems.
  • Strong technical fluency and the ability to understand architecture, system dependencies, quality evidence, operational risk, and engineering trade-offs.
  • Proven ability to design and implement durable operating mechanisms, not merely report status or schedule meetings.
  • Experience defining goals, milestones, ownership, entry and exit criteria, SLAs, risk management, and executive review cadences.
  • Ability to influence technical leaders and teams without direct authority while preserving clear engineering ownership.
  • Strong analytical skills and experience using metrics, dashboards, and qualitative evidence to improve execution and decision quality.
  • Exceptional written and verbal communication, including concise executive synthesis and clear escalation under ambiguity or pressure.

Nice To Haves

  • Experience with AI infrastructure, model delivery, high-performance computing, distributed systems, or hardware/software platforms.
  • Experience supporting feature integration, release qualification, branch stability, developer infrastructure, or production readiness.
  • Familiarity with software quality, E2E testing, CI/CD, observability, incident learning, and reliability mechanisms.
  • Experience coordinating teams with distinct but overlapping ownership boundaries.
  • Experience in a startup or similarly fast-moving, resource-constrained engineering environment.
  • Track record of taking an operating model or cross-team program from zero to one and scaling it as the organization grows.
  • Technical or engineering background sufficient to build credibility with senior engineers and technical leaders.

Responsibilities

  • Define the process for features moving from development and qualification into release integration testing and release qualification, including owners, entry/exit criteria, required evidence, dependencies, and exception paths.
  • Own the operating mechanism for pre-merge and post-merge E2E stability, including publishing health status, tracking SLA breaches, driving triage and escalation, documenting decisions, and closing recurring failure loops.
  • Translate Inference Core priorities into clear goals, milestones, owners, risks, success measures, and review cadences, maintaining a single, dependable view of commitments.
  • Coordinate planning and execution across release integration testing, Core Infrastructure, feature teams, release owners, and partner organizations, synchronizing roadmaps, staffing, and cross-team dependencies.
  • Build and maintain dashboards, SLA reporting, dependency maps, risk registers, decision logs, action tracking, and executive-ready status communications.
  • Drive timely decisions and follow-through on blocked or slipping work, surfacing trade-offs and escalating when issues cannot be resolved at the working level.
  • Continuously improve the operating model using delivery data, retrospectives, recurring failure patterns, stakeholder feedback, and changes in business priorities.

Benefits

  • Job stability with startup vitality
  • Simple, non-corporate work culture that respects individual beliefs
  • Opportunity to build a breakthrough AI platform beyond the constraints of the GPU
  • Opportunity to publish and open source cutting-edge AI research
  • Work on one of the fastest AI supercomputers in the world
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service