Data Center Engineer

INSPYR SolutionsReno, NV
Onsite

About The Position

This role involves collaborating closely with engineering teams to develop and release next-generation products. The Data Center Engineer will manage and maintain a high-performing Compute Farm, ensuring availability targets are met and leading system recovery efforts. The position also requires supporting hardware and software teams, gathering metrics, documenting Standard Operating Procedures (SOPs), and troubleshooting Linux/Windows, hardware, and infrastructure issues. A key aspect of the role is implementing efficiency improvements to enhance availability, throughput, and test accuracy while adhering to SLAs.

Requirements

  • Associate’s or bachelor’s degree in engineering/Technical Major (or equivalent experience).
  • 5+ years with data center technology or large engineering labs.
  • Proficiency in DCIM (Nautobot, etc.) and scripting (shell, Python, Ansible).
  • Working knowledge of protocols/services like TCP/IP, DNS, NFS, SSL, etc.
  • Experience with Windows, Linux, and Mac operating systems.
  • Hands-on experience with PCBs, GPUs, and system deployments.
  • Outstanding communication, both written and verbal.
  • Ability to explain technical concepts to non-technical audiences.
  • Strong problem-solving skills and a collaborative spirit.

Nice To Haves

  • Experience managing HPC clusters using tools like BCM and Slurm.
  • Relevant certifications such as CCNA or equivalent.
  • Strong background in Windows and Linux administration, with an understanding of dense datacenter design, including compute, storage, and networking.
  • Knowledge of DC infrastructure with an emphasis on liquid cooling.
  • Mechanically inclined and comfortable with tools and physical tasks.

Responsibilities

  • Collaborate closely with engineering teams, including system architects, hardware/software engineers, QA, and more, to craft, develop, debug, and release next-generation products.
  • Manage and maintain a high-performing Compute Farm of builders, packagers, testers, and core infrastructure.
  • Ensure availability targets are consistently met and lead system recovery efforts.
  • Support hardware and software teams with hardware and software issues.
  • Gather critical metrics and build Standard Operating Procedures (SOPs) documentation.
  • Problem solve Linux/Windows, hardware, and infrastructure issues alongside engineers and platform operations teams.
  • Implement efficiency improvements to improve availability, throughput, and test accuracy while meeting SLAs and important metrics.

Benefits

  • Equal Employment Opportunities (EEO) to all employees and applicants for employment without regard to race, color, religion, sex, national origin, age, disability, genetics, or any other protected status.
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service