As a Repairs Lead within the Data Center Infrastructure organization, you will define and manage the end-to-end hardware repair program across Anthropic's growing fleet of data centers. You will be accountable for repair turnaround time and the compute returned to service across every site, covering server, GPU/accelerator, network, and optics break-fix, RMA and reverse logistics with OEMs and ODMs, and the spares and repair inventory that keeps repair SLAs achievable. As a subject matter expert in hardware operations, you will develop scalable repair processes and quality targets, set the standards that site operations partners and repair vendors execute against, and turn failure trends into fixes which are driven upstream with engineering and equipment partners. If you are experienced in at-scale datacenter hardware operations, are passionate about HPC data centers, & enjoy working in complex, fast-paced environments, we welcome you to apply.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Senior