Systems SRE Data Center Engineer

Boingo•Frisco, TX
•Hybrid

About The Position

The Systems SRE Datacenter Engineer is responsible for the daily operational maintenance, deployment, and reliability of Boingo’s global production, Staging, QA, Beta, and Dev environments. This hands-on role bridges physical datacenter operations, including racking, stacking, and cabling—with automated cloud-native systems administration. Working as an individual contributor within an SRE and engineering team, you will support HPE Private Cloud Business Edition (PCBE) management platforms, maintain HPE VME and VMware virtualized workloads, monitor high-performance HPE Compute and Storage systems and run automated workflows on Ubuntu Linux to reduce administrative TOIL.

Requirements

  • 3 to 5 years of hands-on experience in systems administration, datacenter engineering, or Site Reliability Engineering (SRE).
  • Solid, working proficiency with Ubuntu OS, RHEL 7+, and standard Linux command-line, system admin tools. Familiarity with Linux networking, systemd, permissions, and package management.
  • Direct experience performing physical datacenter hardware maintenance, including precision racking, server stacking, structured ethernet and fiber cabling, and label management.
  • Hands-on experience operating within HPE Private Cloud Business Edition (PCBE) or HPE GreenLake consoles.
  • Experience supporting virtualized workloads hosted on VMware vSphere or HPE VME environments.
  • Bachelor’s degree preferred.
  • Operational experience deploying, maintaining, and monitoring HPE Compute (ProLiant, Synergy) and HPE Storage systems such as HPE SimpliVity hyperconverged nodes or HPE Nimble Storage arrays.
  • Basic familiarity with HPE iLO remote management APIs.
  • Working experience deploying, troubleshooting, and maintaining containerized applications running on Kubernetes or RKE clusters.
  • Practical ability to execute, maintain, and update Ansible playbooks and Terraform scripts for routine server provisioning and configuration management.
  • Proficiency in Python or Bash to write operational scripts, process logs, and automate repetitive system tasks.
  • Hands-on experience using monitoring tools (Prometheus, Grafana, or InfluxDB) to respond to system alerts, track performance metrics, and troubleshoot outages.
  • Basic understanding of DNS, LDAP, RADIUS, and network routing fundamentals in enterprise environments.
  • Daily familiarity with the Atlassian suite including Jira and Confluence

Responsibilities

  • Execute hardware rack-and-stack installations, structured cabling runs, power connection hookups, and component replacements (RAM, drives, NICs) in local and colocation datacenters.
  • Perform routine patching, software updates, and basic configuration tasks across Ubuntu Linux servers, VMware hosts, and HPE VME hypervisors managed via HPE PCBE.
  • Monitor HPE SimpliVity and Nimble Storage health status; assist with drive replacements, volume allocations, and cluster resource scaling under guidance.
  • Monitor system telemetry dashboards (Grafana, Prometheus), respond to infrastructure alerts, perform initial root-cause analysis, and participate in on-call rotations for production system support.
  • Use Python, Bash, Ansible, and Terraform to automate manual daily tasks, system health checks, and routine deployment workflows.
  • Maintain accurate inventory records, update datacenter rack elevation diagrams, and document technical workflows on the team wiki.
  • Work closely with senior engineers, Network teams, and QA to fulfill ticket requests, troubleshoot environment issues, and stage infrastructure for application deployments.

Benefits

  • health, dental, and vision coverage
  • a 401(k) match
  • unlimited vacation
  • paid parental leave
  • tuition reimbursement
  • cell phone reimbursement
  • pet care benefits
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service