Reliability Engineer

Rubix RecruitingDenver, CO
Onsite

About The Position

You will be working in a fast paced, dynamic development environment designing, developing and delivering our dev and hosting products. You will help maintain a 24x7 uptime on public cloud infrastructure. Be first responder during outages for clients with managed hosting and self-hosting. Contribute to the design and maintenance with regard to logging, networking, monitoring, security and disaster recovery.

Requirements

  • Experience managing production Kubernetes Clusters
  • Fluent in one programming language such as Python, GoLang or Ruby
  • Experience with a blend of knowledge including DevOps, SRE or Systems Operations
  • Experience managing Linux based servers.
  • Understanding of Containers
  • Troubleshooting systems, networks and code
  • Solid knowledge of system performance and monitoring

Nice To Haves

  • Experience with federated Kubernetes clusters
  • Experience with large cloud hosting providers such as AWS, GCP and Azure
  • Experience with Load Balancers
  • Experience with messaging technologies
  • CoreOS (a plus)

Responsibilities

  • Help maintain a 24x7 uptime on public cloud infrastructure.
  • Be first responder during outages for clients with managed hosting and self-hosting.
  • Contribute to the design and maintenance with regard to logging, networking, monitoring, security and disaster recovery.
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service