Critical Infrastructure Engineer

Nebius•Oklahoma, United States, OK
•$49 - $54•Onsite

About The Position

The Infrastructure Engineer is a critical role within Nebius data center operations. A fully proficient Infrastructure Engineer who independently manages and troubleshoots critical power and critical cooling systems across the site and takes ownership of assigned engineering tasks from start to finish. Participates in incident management activities, supports site commissioning and build reviews, and contributes to ensuring all engineering work meets Nebius SLAs and standards. Enforces safe working practices, collaborates with facilities, network, and technician teams, and mentors Associate Infrastructure Engineers toward independent operation.

Requirements

  • Working proficiency across all critical power systems: UPS, PDU, RPP, in-rack busbar, and generator interfacing.
  • Hands-on experience with critical cooling systems: CRAH/CRAC, CDU, RDHx, and water-side infrastructure including leak detection.
  • Competent DCIM/BMS administration: alert management, capacity modeling, and dashboard reporting.
  • Developing knowledge of GB rack liquid cooling technology, DLC manifolds, and heat rejection infrastructure.
  • Ability to participate in and support incident management activities including RCA and CAPA documentation.
  • Strong documentation discipline: change records, maintenance logs, capacity data, and incident write-ups.
  • Clear and confident communication with facilities, network, and vendor teams during changes and incidents.

Responsibilities

  • Monitor and contribute to the tracking of critical environment maintenance and repair for assigned service lines to Nebius SLAs; identify and escalate developing faults before they impact service uptime.
  • Independently operate, monitor, and troubleshoot critical power distribution infrastructure including UPS systems, PDUs, RPPs, and in-rack busbar power systems.
  • Manage and maintain critical cooling infrastructure including CRAHs, CRACs, CDUs, and RDHx; perform capacity checks and leakage inspections.
  • Participate in incident management for infrastructure-impacting events; support root-cause analyses, document findings, and contribute to CAPA execution under the direction of the Senior Infrastructure Engineer.
  • Operate and administer DCIM and BMS platforms; build and maintain dashboards, alerts, capacity reports, and infrastructure records.
  • Support deployment and commissioning of GB-scale liquid-cooled rack infrastructure including direct liquid cooling (DLC) systems, manifolds, and CDU connections.
  • Perform and own preventive maintenance tasks for all critical power and critical cooling systems; maintain accurate maintenance logs and compliance records.
  • Participate in site reviews, design reviews, and commissioning activities; prepare and review technical reports to document findings and communicate results.
  • Enforce Nebius critical power and critical cooling safety procedures; act as safety lead for engineering work orders on live infrastructure.
  • Collaborate with facilities, network, and technician teams on infrastructure changes, capacity expansions, and major deployments.
  • Support adherence to Nebius standards and policies through documentation review and active participation in commissioning and design activities.
  • Contribute to vendor and contractor coordination by supporting scheduling, site access, and execution of work per Nebius expectations and safe-working practices.
  • Actively mentors and trains Associate Associate Infrastructure Engineers; guides them through systems operation, safe working practices, and structured skill development toward independent operation.

Benefits

  • Competitive compensation
  • Career growth and learning opportunities
  • Flexibility and ownership
  • Collaborative and innovative culture
  • Opportunity to work on impactful AI projects
  • International environment and talented teams
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service