About The Position

You'll own Luma's global compute footprint end to end — capacity strategy, multi-million-dollar capital allocation, and systems architecture — making sure research and robotics teams have the runway to ship frontier world models. As a member of the executive team, you're the single person turning capital into capability. The role spans macro capacity strategy, vendor negotiation, and top-tier systems architecture, and it directs the platform org. It fits a leader who's operated 10k+ accelerator environments and is fluent in both cluster topology and the economics of training. If you're looking for a purely technical or purely strategic seat, this is deliberately both.

Requirements

  • 10+ years of engineering leadership in large-scale distributed systems, infrastructure, or technical supply chain, with a track record leading compute platform strategy at a frontier AI lab, hyperscaler, or major autonomy program.
  • Deep technical and commercial fluency in cluster topology, high-speed interconnects (InfiniBand/RoCE), large-scale data systems, and the economics of distributed training.
  • Direct operational oversight of 10k+ accelerator environments in production.

Nice To Haves

  • Experience orchestrating capital or infrastructure for training runs at the 100B-parameter or 100k-GPU-day scale.
  • Familiarity with the capacity and latency demands of edge-to-cloud inference and real-time autonomous systems.

Responsibilities

  • Architect multi-year compute strategy: capacity planning, global vendor and cloud partnerships, on-prem vs cloud mix, accelerator supply-chain roadmaps, and custom-silicon evaluation.
  • Provide strategic leadership to infrastructure, distributed systems, and datacenter operations teams.
  • Maximize fleet utilization, targeting more than 50% Model Flops Utilization on flagship training runs.
  • Negotiate, secure, and operate the largest-scale capital deployments, partnering with Finance on unit economics and risk.
  • Unify global capacity so world-model training, simulation, and on-robot inference share a single elastic fleet.
  • Act as the principal executive interface to NVIDIA, AMD, hyperscalers, and frontier silicon vendors.
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service