Crusoe Cloud is seeking a Senior Network Production Operations Engineer to support production reliability across our global network infrastructure, including edge, backbone, data center fabric, and GPU cluster interconnects. This is a hands-on production role focused on incident response, root cause analysis, and automation-driven operational execution that keeps our hyperscale AI infrastructure running at scale. Your work will directly affect the availability of AI workloads running across thousands of GPUs worldwide. The ideal candidate is an experienced network engineer with solid operational experience in large-scale environments who thrives in high-pressure situations and takes pride in keeping systems healthy. You'll build automation to reduce operational toil, execute on established SLIs and SLOs, contribute to observability tooling, and serve as a responder during high-severity network events.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Senior