​​High Performance Computing (HPC) Systems Administrator​

LeidosSan Diego, CA
$107,900 - $195,050Onsite

About The Position

The Leidos Defense Sector currently has an opening for a High Performance Computing (HPC) Systems Administrator to work in our San Diego, California office. This is an exciting opportunity to use your experience supporting the RIPTIDE environment. The High Performance Computing (HPC) Systems Administrator will leverage technical expertise to collaborate with the existing HPC team, architects, and engineers in supporting the computational needs of the Defense Sector. Reporting to the IT Security Engineering Manager, the Systems Administrator will ensure that the HPC cluster is maintained at operational levels and support system users. Primary responsibilities include system administration, configuration, monitoring, troubleshooting, software maintenance, installations, SLURM management and job scheduling system for the Linux clusters, and optimization of advanced computing systems and related infrastructure.

Requirements

  • Bachelor’s degree in computer science or a related STEM field, combined with 3-5+ years of dedicated HPC or large-scale Linux system administration experience.
  • Advanced knowledge of SLURM Workload Manager configurations, including accounting, resource limits, fair-share scheduling, and node state management.
  • Hands-on experience with automation frameworks like ansible.
  • Proficiency in scripting languages such as Bash, Python, or Perl for system administration and automation.
  • Strong troubleshooting skills across systems, hardware, network, and application layers.
  • Familiarity with software installation, configuration, and maintenance.
  • Knowledge of security best practices, including the handling of protected data.
  • Excellent communication and documentation skills. Qualified candidates must be able to effectively communicate with all levels of the organization.

Nice To Haves

  • Red Hat Certified System Administrator and/or Engineer Certification.
  • Experience in integrating HPC resources.

Responsibilities

  • Deploy, administer, and monitor Linux-based HPC clusters and storage systems.
  • Install, maintain, and upgrade software, libraries, and licensed applications.
  • Troubleshoot and resolve system, hardware, and application issues across HPC environments.
  • Develop and maintain scripts for system administration, monitoring, reporting, and data pipeline optimization.
  • Provide technical support to system users; assist with project execution, benchmarking, and resource planning.
  • Research, deploy, and optimize resource management and scheduling systems.
  • Evaluate and manage a helpdesk ticketing system; ensure timely resolution of user-submitted issues.
  • Implement and manage HPC security infrastructure; ensure compliance with data handling policies and best practices.
  • Create and maintain technical documentation and user guides.
  • Consult system users to assess computational needs.
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service