Linux Cluster Hardware Administrator

University of Texas at AustinAustin, TX
$65,000Onsite

About The Position

The Texas Advanced Computing Center (TACC) at The University of Texas at Austin is one of the leading supercomputing centers in the world, supporting advances in computational research by thousands of researchers and students. TACC staff help researchers and educators use advanced computing, visualization, and storage technologies effectively, and conduct research and development to make these technologies more powerful, more reliable, and easier to use. TACC staff also educate and train the next generation of researchers, empowering them to make discoveries that advance knowledge and change the world. The Large Scale Systems group at the Texas Advanced Computing Center is seeking a hardware-focused Systems Administrator with extensive deployment experience of OmniPath 400Gbp/s and XDR Infiniband fabrics, in a large scale academic high performance computing environment. Demonstrable experience with hardware upgrades from OmniPath 100Gbp/s to OmniPath 400Gbp/s is required. This position will require interaction between various vendors and on campus teams in a high performance computing environment, spanning two physical locations that include Petascale HPC /storage systems, and other high-speed networking infrastructure. Candidates will need to upload a resume, letter of interest, and the names of three references to apply for this position. UT Austin offers a competitive benefits package that includes: 100% employer-paid basic medical coverage Retirement contributions Paid vacation and sick time Paid holidays Please visit our Human Resources (HR) website to learn more about the total benefits offered. If you are not sure that you’re 100% qualified, but up for the challenge – we want you to apply. We believe skills are transferable and passion for our mission goes a long way. Must be eligible to work in the US on a full-time basis for any employer without sponsorship.

Requirements

  • 5+ years of hands’ on experience with trouble-shooting and repair of telecommunications equipment and computer hardware
  • Experience with OmniPath 400Gbp/s in an HPC environment
  • Experience with XDR Infiniband in an HPC environment
  • Experience with Linux systems administration tools
  • Experience with server hardware and physical repairs of high-density HPC systems
  • Experience with physical deployments including cabling, power distribution, and familiarity of advanced datacenter cooling technologies (DLC and immersion cooling)
  • Ability and desire to learn new concepts and software tools
  • Excellent verbal/written communication skills.
  • Must be eligible to work in the US on a full-time basis for any employer without sponsorship.

Nice To Haves

  • Experience in physical installation of large-scale clusters
  • Experience with clusters that have multiple interconnect fabrics: multiple Infiniband technologies/OmniPath 100Gbp/s & OmniPath 400Gbp/s
  • Working knowledge of systems provisioning techniques and tools
  • Basic knowledge with one or more scripting languages, (Python, Perl, Bash)
  • Working knowledge of monitoring software such as Nagios or Zabbix

Responsibilities

  • Diagnose and repair Linux cluster hardware components: Standard datacenter server components including all air-cooled components, Direct-Liquid-Cooling (DLC) infrastructure components including cold-plates and coolant distribution units (CDU)s, Coolant-Oil-Immersion-Cooling (COIC) subsystem components including CDUs, Petabyte-scale storage hardware components
  • Maintain up to date catalog of all work performed
  • Deploy new systems and subsystems: Deployment and management of high-speed networking fabrics, Optimizing rack positioning and internal layout to ensure power/thermal efficiency

Benefits

  • 100% employer-paid basic medical coverage
  • Retirement contributions
  • Paid vacation and sick time
  • Paid holidays
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service