Systems Administrator, GPU/AI Infrastructure

Lenovo•Morrisville, NC
•Onsite

About The Position

Lenovo seeks a Systems Administrator to own end-to-end deployment, support, and administration of the consolidated Morrisville AI lab. This role covers network, compute, storage, operating system, and container orchestration (Docker and Kubernetes) layers, as well as the NVIDIA technology stack across over 50 servers. It is a hands-on infrastructure role where the individual will perform day-to-day administration, support, and engineering tasks, including rack-and-stack and smart hands. The Systems Administrator will be the sole dedicated support resource for the Lab Operations Manager, ensuring the operational status of the lab's significant capital investment. This role involves collaboration with a Network Engineer and a Junior Systems Administrator, providing full-stack administration to connect their specialized work. Additionally, the role will offer secondary technical support for equipment in Bangalore and validate lab configurations against NVIDIA reference architecture certification standards. The position is part of Lenovo’s Hybrid Cloud and AI Infrastructure Services (HCAIS) Lab Operations organization.

Requirements

  • 2+ years of systems administration experience with hands-on breadth across network, compute, storage, and operating system layers, plus a GPU/AI infrastructure specialization.
  • Production Kubernetes Administration: demonstrated experience with cluster operations, RBAC, cluster networking, and version upgrades in a production environment.
  • Docker: hands-on experience administering Docker container runtime environments.
  • Hands-on NVIDIA administration: demonstrated experience administering NVIDIA drivers and the CUDA toolkit.
  • InfiniBand Fabric: working familiarity with InfiniBand fabric in a multi-node GPU environment.
  • GPU Orchestration: experience with Run:AI, ClearML, or a comparable GPU orchestration and MLOps toolset.
  • Physical Infrastructure: comfortable with hands-on hardware work, including rack-and-stack, structured cabling, and smart hands, in addition to higher-level administration.
  • Incident Response: able to own after-hours incident response for a shared, multi-team infrastructure environment.
  • Cross-Team Coordination: comfortable supporting multiple HC/AI teams across a shared lab environment with competing priorities, and coordinating day to day with the Network Engineer and Junior Systems Administrator roles on adjacent, overlapping infrastructure.

Nice To Haves

  • Kubernetes Certification: Certified Kubernetes Administrator (CKA) or equivalent.
  • Multi-Tenant Lab Experience: experience supporting a multi-tenant shared lab environment serving distributed teams.
  • Familiarity with NVIDIA certification and reference architecture (Lenovo Validated Design) processes.
  • Experience with DCIM, IPAM, or monitoring tooling comparable to the lab’s stack (Hyperview-class DCIM, BlueCat / Infoblox-class IPAM, NVIDIA DCGM monitoring).

Responsibilities

  • Deploy, support, and administer lab networking, coordinating with the Network Engineer for switch and fabric configuration.
  • Deploy, support, and administer physical and virtual compute infrastructure.
  • Deploy, support, and administer lab storage systems across block and file tiers.
  • Install, patch, and administer operating systems across lab infrastructure.
  • Deploy and administer Docker container runtime environments.
  • Own full production Kubernetes administration, including cluster operations, role-based access control (RBAC), cluster networking, and version upgrades.
  • Administer NVIDIA drivers, CUDA toolkit, InfiniBand fabric, Run:AI orchestration, and ClearML across the GPU infrastructure.
  • Validate lab configurations against NVIDIA reference architecture (Lenovo Validated Design) certification requirements.
  • Provide day-to-day administration, support, and engineering functions for the network, compute, and storage layers in the physical lab.
  • Perform rack-and-stack, structured cabling, and smart-hands support for lab hardware.
  • Provide secondary technical support for NVIDIA technology stack capital equipment hosted in Bangalore.
  • Coordinate with the adjacent ISG AI COE lab (Tech Marketing, customer proof-of-concepts) and the Morrisville Executive Briefing Center.
  • Maintain a direct working relationship with NVIDIA field contacts.
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service