Cerebras is building and operating some of the world's most advanced AI infrastructure. As our infrastructure footprint expands across cloud, colocation facilities, and customer environments, we need a scalable operating model that provides continuous visibility into production infrastructure and ensures issues are rapidly identified, prioritized, and resolved. We are seeking a Senior Manager, Production & Fleet Operations to lead the day-to-day operational management of Cerebras-managed production infrastructure. Reporting to the Director of Central Operations, this leader will establish and operate the mechanisms required to understand fleet health, coordinate production response, manage maintenance and repair activities, and ensure infrastructure is safely and efficiently returned to service when failures occur. This role will work closely with SiteOps/DC Ops, Reliability & Incident Management, Service Ops & Enablement, Global Service Logistics & Inventory, Build & Deploy, and Engineering. The mission is to operate the Cerebras production fleet safely, reliably, and consistently, providing continuous visibility into infrastructure health, rapid restoration when failures occur, and an increasingly standardized and automated operating model as the fleet scales.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Senior
Education Level
No Education Listed