Senior HPC Support Engineer, InfiniBand - NVLink

NVIDIADurham, NC
$108,000 - $207,000

About The Position

NVIDIA is seeking a highly motivated Senior HPC Support Engineer focused on InfiniBand and NVLink technology, with a passion for data center and networking technologies. The role involves providing comprehensive solutions for sophisticated installations, maintenance, and operations for groundbreaking networking products. As a primary point of contact for customers, you will assist with technical questions, debugging, and resolving issues. As part of the NVIDIA Experience (NVEX) Global Technical Support team, you will be a conscientious, proficient communicator, taking ownership in resolving issues and ensuring high customer satisfaction. A significant part of the role includes regular collaboration with Engineering, Marketing, and Support teams on technical matters.

Requirements

  • 5+ years in providing in-depth Customer Support and debugging experience focused on large-scale networking and AI Infrastructure environments and products such as InfiniBand, Ethernet, and GPU including infrastructure performance.
  • Strong interpersonal skills and ability to prioritize/multi-task easily with limited supervision.
  • Shown use of established Agentic AI technologies (Claude, Codex, Cursor) in day-to-day job responsibilities.
  • Excellent verbal and written English skills.
  • An academic degree from an accredited university or college in Networking, Computer Science/Engineering, or Electrical/IT (or equivalent experience).
  • Intellectual curiosity, positive attitude, flexibility, analytical ability, self-motivation, and team-oriented including professional-level communication skills, interpersonal skills with a passion to solve problems.
  • Profound knowledge and experience (solving) in Networking Technology, protocols and routing including IP, L2 and L3 on a CCNP/CompTIA Networking+ and Cloud+ level.
  • Able to debug networking protocols using tools such as TCPDUMP and Wireshark or similar packet generation and analysis tools.
  • Configuration and operational expertise with traditional network switch/router and Open platforms.
  • Linux OS System Administration, Networking and Performance on a LFCS/RHCSA level.
  • Containerized solutions experience on a level of DCA and/or CKA, Virtualization and (KVM/ESXi) and Cloud Infrastructure (AWS/OCI) Technologies.

Nice To Haves

  • Knowledge and working experience with InfiniBand, RDMA/RoCEv2 and GPU Technology.
  • Clustering or HPC Data-Center technologies including Upper Layer Protocols (i.e., NCCL, MPI, Slurm/SchedMD).
  • Shell scripting (Bash/Python).
  • Linux, Networking and NVIDIA AI Infrastructure and Operations Certifications such as CCNP, CCIE, JNCIE-DC/ENT, RHCE, LFCS, NCP-AII/AIO/AIN.

Responsibilities

  • Resolve sophisticated customer concerns and technical issues through meticulous research, reproduction, and problem-solving for customers installing products and supporting systems using Linux Operating Systems (Multi-distro), with a focus on NVIDIA InfiniBand, NVLink, GPU Technology, and End-to-End Solutions.
  • Respond to customer product support inquiries via telephone, email, or conference calls.
  • Resolve customer issues during installation, operation, maintenance, product application, or interoperability with other vendors.
  • Participate in multi-functional team meetings and provide feedback to engineering and marketing regarding product requirements, customer experience, and support tools.
  • As a technical resource, develop, redefine, and document standard methodologies for internal teams (Support/R&D) for support process and improvements.

Benefits

  • Highly competitive salaries
  • Comprehensive benefits package
  • Equity
  • Benefits
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service