Cleared Senior Site Reliability Engineer

GallatinAustin, TX
$80,000 - $210,000Hybrid

About The Position

Gallatin is looking for a Site Reliability Engineer to keep our production systems running with the reliability our national security customers require. You'll work at the intersection of infrastructure, automation, and mission-critical uptime — building and operating the systems that turn logistics data into decisions in real time, including in classified and disconnected environments. This is a hands-on role for an engineer who treats reliability as a product: someone who instruments before things break, automates the toil away, and owns incidents from detection through postmortem.

Requirements

  • Active Secret clearance required; willingness and eligibility to obtain a higher-level clearance if needed.
  • 3–5 years of experience in a site reliability engineering, DevOps, or production infrastructure role.
  • Strong technical background in Linux systems, networking, and cloud infrastructure (AWS, Azure, or GovCloud equivalents).
  • Hands-on experience with infrastructure-as-code (Terraform, Ansible, or similar) and container orchestration (Kubernetes, Docker).
  • Experience building and maintaining CI/CD pipelines and automated deployment systems.
  • Familiarity with monitoring and observability tooling (Prometheus, Grafana, Datadog, ELK, or similar).
  • Track record of owning production incidents end-to-end, from detection through resolution and postmortem.
  • Comfort operating in classified, air-gapped, or otherwise network-constrained environments.
  • Clear written and verbal communicator who can work directly with both engineers and government stakeholders.
  • Willing to travel >50% of the time to customer and government sites.
  • U.S. citizenship is a requirement for all positions at Gallatin.

Nice To Haves

  • Hands-on experience with Microsoft Azure, including Azure Government (GCC High) or Azure Government Secret / IL5–IL6 environments.
  • Experience achieving or maintaining an Authority to Operate (ATO) under RMF in a DoD or federal environment.
  • Experience standing up or operating in multi-cloud environments spanning AWS and Azure.
  • Familiarity with edge or disconnected/degraded/intermittent/limited (DDIL) deployment patterns for classified or tactical environments.
  • Prior experience at an early-stage startup or other fast-moving, high-ownership engineering team.
  • Background in logistics, supply chain, or defense sustainment systems.

Responsibilities

  • Own the reliability, availability, and performance of production systems supporting Gallatin's logistics decision platform, including services deployed in classified and air-gapped environments.
  • Build and maintain monitoring, alerting, and observability pipelines that surface problems before customers do.
  • Lead incident response for production issues, driving triage and resolution and running blameless postmortems that turn into concrete engineering fixes.
  • Design and operate CI/CD pipelines and infrastructure-as-code that let engineering ship safely and often, across both cloud and classified network environments.
  • Automate manual operational work (deployments, scaling, failover, credential rotation) to reduce toil and remove single points of failure.
  • Harden systems to meet DoD security and compliance requirements (e.g., RMF, STIGs, ATO processes) without slowing down delivery.
  • Partner with software engineers to define SLOs/SLIs and build reliability into services from design through deployment.
  • Work directly with government customers and field teams to understand mission environments and translate operational constraints into system requirements.
  • Document runbooks, architecture decisions, and operational procedures so the systems you build can be run by the whole team, not just you.

Benefits

  • generous equity grant
  • full healthcare coverage
  • 401k
  • unlimited PTO
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service