Senior DevOps Engineer

1X•San Carlos, CA
•$190,000 - $230,000•Onsite

About The Position

We're building humanoid robots that work in home - doing the chores, handling the tasks, and giving people their time back. The Cloud Infrastructure team owns the platform every other team at 1X builds on: our Kubernetes clusters, AWS accounts, CI systems, identity and access layer, and the GPU compute our AI, robotics, and fleet teams depend on every day. We're a small team with an outsized blast radius - there's no SRE org or ticket queue between us and the engineers we serve, so when a training run needs a thousand GPUs by Monday, a cluster goes down at 2am, or a robot in someone's home needs a secure path back to our cloud, it lands with us. The mission of this role is to make 1X's cloud infrastructure reliable, secure, and self-service enough that every team building NEO can move as fast as the hardware allows. You'll own the reliability and scalability of our cloud infrastructure end-to-end, from architecture decisions through incident response, and build the automation and self-service tooling that lets the broader engineering team handle common requests safely on their own, without piling up technical debt. You'll also be a steady hand for day-to-day urgent requests and team support when the platform needs it most.

Requirements

  • Proven track record of 5+ years in a DevOps, SRE, or infrastructure engineering role.
  • Experience with hands-on production operation of a major cloud provider or large-scale on-prem infrastructure using modern tooling.
  • Experience with service-to-service communication, specifically identity and auth.
  • Demonstrated ability to write production code in Python and/or Go.
  • Experience with Terraform.

Nice To Haves

  • Experience in improving incident response processes.
  • Experience with AWS, Cloudflare, and/or GPU cloud providers.
  • Experience in infrastructure security.
  • Experience with chaos engineering and improving cloud reliability under adverse conditions.
  • Experience in improving CI as a platform.

Responsibilities

  • Design and maintain infrastructure-as-code systems that make provisioning and scaling reliable and easy.
  • Build self-service tooling that lets engineers safely handle common cloud requests without creating technical debt.
  • Drive incident response with clear communication and urgency, and lead post-incident remediations so we systematically improve.
  • Partner with the security team to strengthen cloud infrastructure security and reliability.

Benefits

  • Comprehensive medical, dental, and vision coverage
  • Generous paid time off, company holidays, and parental leave
  • 401(k) plan with company match (100% on the first 3% of contributions, 50% on the next 2%)
  • Flexible Spending Accounts (FSA) and Health Savings Accounts (HSA) options
  • Commuter benefits (transit and parking)
  • Short-term and long-term disability, and life insurance
  • Employee Assistance Program (EAP) for mental health, financial, and personal support
  • Onsite snacks and catered lunches
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service