AI Factory CPU focused Solutions Architect

NVIDIASanta Clara, CA
$184,000 - $356,500Onsite

About The Position

We are looking for a hardworking Solution Architect with experience in designing, building, and maintaining large-scale HPC and AI infrastructure to join our team at NVIDIA. As Solution Architects, we are actively helping make AI Factories a reality. Working closely with customers and partners to address unsolved problems in the industry, our team helps to deploy and operationalize AI solutions at scale. Our day-to-day work involves helping our partners be successful in their adoption of end-to-end AI solutions using NVIDIA's compute, networking, and software stacks. For this particular role, that means having a deep technical understanding of NVIDIA Reference Architectures, and using that understanding to enable customers adopting our CPU-based solutions as part of the overall NVIDIA AI Factory. This is a multi-faceted role necessitating being comfortable working on not just hardware and software elements, but also the larger AI workflow and operationalization of large scale compute resources. We succeed when we help our customers overcome barriers to adopting our best known methods. As the technical leader for the CPU components within the NVIDIA AI Factory, you will play an instrumental role in driving that success. As a team, we also excel at sharing knowledge with our colleagues, whether it's delivering demos, assisting with proof-of-concepts, or writing papers and developer blogs. By collaborating with executives and engineering, we tackle sophisticated problems and help bring NVIDIA's premiere technologies to life. Our mission is to solve the problems that nobody else has solved yet, and we need someone to be an instrumental part of that!

Requirements

  • Experience with defining, deploying, and testing large scale reference architectures for High Performance Computing and AI.
  • A track record of defining and using MLOps and AI workflow tools and processes.
  • 6 or more years of hands-on expertise with modern data center architectures and interaction between CPUs, GPUs, and networking.
  • Strong foundational expertise and a BS, MS, or equivalent experience in Engineering.
  • Strong analytical and problem-solving skills.
  • Ability to articulate technical knowledge to others.
  • Ability to multitask efficiently in a multifaceted environment.
  • Experienced with organizing, presenting, and discussing technical materials with groups of varying technical capability.
  • Flexibility to adapt in fluid situations, especially with partners or customers.
  • Comfortable with occasional travel to customer sites.

Nice To Haves

  • Hands-on experience with Arm-based server processors and the Arm software ecosystem.
  • Proficiency with tooling, automation, and performance testing for large-scale clusters, preferably using AI tools.
  • Deep understanding of Agentic AI and inference workflows.
  • Experience building, using, and explaining reinforcement learning.
  • Willingness and ability to learn quickly as we address sophisticated problems.
  • Understanding of how all elements of the AI Factory interact with each other.

Responsibilities

  • Designing, building, and maintaining large-scale HPC and AI infrastructure.
  • Helping partners be successful in their adoption of end-to-end AI solutions using NVIDIA's compute, networking, and software stacks.
  • Enabling customers adopting CPU-based solutions as part of the overall NVIDIA AI Factory.
  • Working on hardware and software elements, the larger AI workflow, and operationalization of large scale compute resources.
  • Helping customers overcome barriers to adopting best known methods.
  • Delivering demos, assisting with proof-of-concepts, or writing papers and developer blogs.
  • Collaborating with executives and engineering to tackle sophisticated problems and help bring NVIDIA's premiere technologies to life.

Benefits

  • Equity
  • Benefits
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service