About The Position

Mistral is building a sovereign AI cloud for Europe. Our customers in finance, healthcare, and the public sector run frontier models on European-jurisdiction infrastructure, and the Cloud Platform team builds the systems that make that possible — compute orchestration, Kubernetes, bare metal, storage, networking, and the datacenter fleet that runs it all. European Compute Units let enterprises commit to multi-year capacity on that infrastructure, and we are scaling toward a gigawatt of compute across Europe by 2030. You will design and ship the layer every sovereign AI deployment depends on and shape the team building it.

Requirements

  • Automated bare-metal fleets — you have written the software that turns racks into clusters. If you have done it by hand, you resented it and scripted it.
  • Hardware fluency: servers, GPUs, NICs, switches, firmware — you can read a parts list and a topology diagram.
  • 8+ years building backend or infrastructure systems, with at least 2 years leading engineers — real people leadership: you have hired, coached, and grown engineers, not just set technical direction.
  • Deep Linux: you are comfortable well below the application layer.
  • You have shipped production systems in Go. Not dabbled. Shipped.
  • Willing to travel to datacenter sites when the work demands it.
  • A low-ego, team-first mindset. We care less about your exact domain and more about whether you are the kind of technical force other engineers want to work alongside.

Responsibilities

  • Build the delivery pipeline in Go: provisioning, configuration, validation, burn-in, network bring-up, software stack install.
  • Build the team from scratch: make the first hires, onboard and coach them, own their performance and growth paths — and set the tooling standards and the playbook every future site reuses.
  • Own time-to-acceptance and drive it down relentlessly — measure it, break it down, kill the slowest step first.
  • Work shoulder-to-shoulder with Hardware Design and Fleet Health — reliability starts at delivery.
  • Handle GPU fleets at scale: H100 and GB200-class nodes, InfiniBand and RoCE fabrics, and the failure modes they bring.
  • Partner with Product Managers and the datacenter org to plan delivery waves against capacity commitments.

Benefits

  • healthcare coverage
  • parental leave
  • retirement plans
  • relocation support
  • wellness programs
  • meal and transportation allowances
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service