Technical Project Manager, GPU Infrastructure Deployment

5C Data Centers USA, Inc.Springfield, OH
$155,000 - $175,000Remote

About The Position

We are seeking a Technical Project Manager (TPM) to drive the planning, coordination, and execution of large-scale GPU cluster infrastructure deployment programs. This role partners across multiple engineering teams to manage project schedules, track milestones, resolve blockers, and communicate status across technical teams, datacenter operations, engineering, and external vendors. The TPM operates at the intersection of program leadership and hands-on technical deployment, ensuring that end-to-end GPU cluster buildouts - spanning compute, high-speed networking fabrics, high-performance storage, power, cooling, and automation - are delivered on time, within budget, and to operational readiness standards. This includes coordinating mechanical, electrical, and plumbing (MEP) readiness with facilities engineering and datacenter operations to ensure power capacity, cooling distribution, and physical infrastructure are in place ahead of rack-and-stack. The ideal candidate brings strong technical project management expertise in hyperscale or cloud infrastructure environments, combined with a working understanding of AI/GPU deployment lifecycles and cross-functional program execution.

Requirements

  • Lead end-to-end project management for large-scale AI/GPU cluster deployments, including multi-rack GPU compute platforms (NVIDIA DGX and similar), InfiniBand and Ethernet GPU fabrics, and high-performance storage environments (e.g., VAST Data).
  • Partner with engineering teams and managers to establish and drive consistent project tracking, milestone reporting, and status updates across the team and systems (e.g., Jira, Confluence).
  • Coordinate procurement, rack-and-stack sequencing, cabling schedules, network deployment timelines, burn-in testing, cluster validation, and operational handoff.
  • Coordinate with facilities engineering and datacenter operations to align MEP readiness (power distribution, cooling capacity, floor layout, containment) with deployment schedules.
  • Communicate project status, risks, milestones, and dependencies to stakeholders at all levels including executive leadership.
  • Prepare and present regular program reviews, steering committee updates, and ad-hoc project analyses.
  • Maintain centralized project documentation and ensure consistent, accurate data.
  • Contribute to improving and documenting repeatable deployment methodologies, scalable operational standards, and project management best practices.
  • Track deployment KPIs (schedule variance, budget adherence, quality metrics) and drive continuous improvement through data-driven retrospectives.

Nice To Haves

  • Experience with hyperscale cloud providers, large-scale data center deployments, or AI infrastructure programs.
  • Familiarity with: Infrastructure-as-Code and configuration management (Ansible)
  • Familiarity with: Python, Shell, and SQL for infrastructure automation and diagnostics
  • Familiarity with: DCIM, BMS, and observability platforms for liquid-cooled environments
  • Familiarity with: NVIDIA/Mellanox networking platforms and NVLink/NVSwitch technologies
  • Experience coordinating with MEP engineering teams on electrical and cooling infrastructure scheduling for large-scale deployments.
  • Experience with procurement processes in infrastructure delivery contexts.

Responsibilities

  • Lead end-to-end project management for large-scale AI/GPU cluster deployments, including multi-rack GPU compute platforms (NVIDIA DGX and similar), InfiniBand and Ethernet GPU fabrics, and high-performance storage environments (e.g., VAST Data).
  • Partner with engineering teams and managers to establish and drive consistent project tracking, milestone reporting, and status updates across the team and systems (e.g., Jira, Confluence).
  • Coordinate procurement, rack-and-stack sequencing, cabling schedules, network deployment timelines, burn-in testing, cluster validation, and operational handoff.
  • Coordinate with facilities engineering and datacenter operations to align MEP readiness (power distribution, cooling capacity, floor layout, containment) with deployment schedules.
  • Communicate project status, risks, milestones, and dependencies to stakeholders at all levels including executive leadership.
  • Prepare and present regular program reviews, steering committee updates, and ad-hoc project analyses.
  • Maintain centralized project documentation and ensure consistent, accurate data.
  • Contribute to improving and documenting repeatable deployment methodologies, scalable operational standards, and project management best practices.
  • Track deployment KPIs (schedule variance, budget adherence, quality metrics) and drive continuous improvement through data-driven retrospectives.

Benefits

  • Competitive pay
  • Career Growth
  • Industry Leadership
  • Entrepreneurial Culture
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service