The HPC/AI (High performance Computing and Artificial Intelligence) team is on a mission to build the next-generation distributed AI supercomputer, enabling breakthroughs in artificial intelligence by delivering unmatched computational power, scalability and reliability. We design and develop cutting-edge infrastructure that supports high-performance AI model training at scale, laying the foundation for innovations that redefine what AI can achieve. You will help in the delivery and operations of the network infrastructure and tooling that powers the next generation of large-scale AI and HPC networking systems. In this role, you will manage AI network fabric that are critical for achieving ultra-low latency, high throughput, strong reliability and petabyte-scale efficiency in distributed AI workloads. As a Senior Cloud Network Engineer on the HPC & AI Infrastructure team, you’ll work at the intersection of AI supercomputing and large-scale networking, shaping how advanced AI models are trained and deployed in the cloud. Your contributions will directly impact the reliability and performance of massive distributed clusters, leveraging high-speed fabrics (e.g., InfiniBand, RoCE) and accelerated compute platforms (e.g., NVIDIA, AMD GPUs). Microsoft’s mission is to empower every person and every organization on the planet to achieve more. As employees we come together with a growth mindset, innovate to empower others, and collaborate to realize our shared goals. Each day we build on our values of respect, integrity, and accountability to create a culture of inclusion where everyone can thrive at work and beyond.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Senior
Education Level
No Education Listed