Senior DevOps Engineer

1X•San Carlos, CA
•Onsite

About The Position

1X is building humanoid robots for home use, aiming to give people their time back by handling chores and tasks. This involves solving complex challenges in robotics, AI, and manufacturing simultaneously and at scale, with a focus on safety for home environments. The company has been developing its flagship product, NEO, since 2014 and is now focused on shipping it. 1X is seeking individuals inspired by this mission who want to build something that will genuinely change how humans spend their time, safely creating abundance for all. The Cloud Infrastructure team is responsible for the platform on which all other teams at 1X build, including Kubernetes clusters, AWS accounts, CI systems, identity and access management, and the GPU compute resources essential for AI, robotics, and fleet teams. This small team has a significant impact, directly supporting engineers without an intervening SRE organization or ticket queue, and managing critical situations like urgent GPU needs, cluster outages, or secure cloud connectivity for robots in homes. The mission of this role is to ensure 1X's cloud infrastructure is reliable, secure, and self-service, enabling all teams building NEO to operate at maximum speed. The Senior DevOps Engineer will own the end-to-end reliability and scalability of the cloud infrastructure, from architectural decisions to incident response. They will also develop automation and self-service tools to empower the broader engineering team to handle common requests independently, minimizing technical debt. Additionally, the role involves providing steady support for day-to-day urgent requests and team needs when the platform requires it.

Requirements

  • Proven track record of 5+ years in a DevOps, SRE, or infrastructure engineering role.
  • Experience with hands-on production operation of a major cloud provider or large-scale on-prem infrastructure using modern tooling.
  • Experience with service-to-service communication, specifically identity and auth.
  • Demonstrated ability to write production code in Python and/or Go.
  • Experience with Terraform.
  • Deep understanding of distributed systems, networking, and cloud architecture trade-offs.
  • Strong infrastructure-as-code mindset.
  • Pragmatic about balancing speed against long-term architecture.
  • Calm, clear communicator who brings urgency to incident response.
  • Bias toward automation and self-service tooling over manual toil.
  • Customer-focused attitude that shows up in how you support urgent requests from the team.

Nice To Haves

  • Experience in improving incident response processes.
  • Experience with AWS, Cloudflare, and/or GPU cloud providers.
  • Experience in infrastructure security.
  • Experience with chaos engineering and improving cloud reliability under adverse conditions.
  • Experience in improving CI as a platform.

Responsibilities

  • Design and maintain infrastructure-as-code systems for reliable and easy provisioning and scaling.
  • Build self-service tooling to allow engineers to safely handle common cloud requests without creating technical debt.
  • Drive incident response with clear communication and urgency, and lead post-incident remediations for systematic improvement.
  • Partner with the security team to enhance cloud infrastructure security and reliability.

Benefits

  • Comprehensive medical, dental, and vision coverage
  • Generous paid time off, company holidays, and parental leave
  • 401(k) plan with company match (100% on the first 3% of contributions, 50% on the next 2%)
  • Flexible Spending Accounts (FSA) and Health Savings Accounts (HSA) options
  • Commuter benefits (transit and parking)
  • Short-term and long-term disability, and life insurance
  • Employee Assistance Program (EAP) for mental health, financial, and personal support
  • Onsite snacks and catered lunches
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service