Senior Site Reliability Engineer II

Juniper Square,
$165,000 - $195,000Remote

About The Position

We're looking for a Senior Site Reliability Engineer II to help us scale our infrastructure and reliability practices as we grow our engineering org by 90+ people this year. You'll take ownership of reliability and observability for your team's systems, work hands-on across Kubernetes, AWS, and CI/CD, and partner with 16+ engineering teams to roll out standards and automation that reduce friction org-wide. This role is for someone who's proactive by nature — you see a problem and go fix it without waiting to be asked. You're comfortable working across teams with different workflows and priorities, and you're already using AI tools as part of how you build, debug, and ship. You won't be working in a vacuum: you'll partner with a Staff-level architect and engineering leadership on the biggest calls, while owning execution and cross-team rollout yourself.

Requirements

  • 8-10 years of experience as a Senior SRE, DevOps, or Infrastructure Engineer
  • Background at smaller-to-mid-size, high-growth SaaS companies (Series A/B through C/D) — you know how to move fast and own ambiguity, not navigate a large, slow-moving org
  • Deep hands-on experience with Kubernetes (EKS or ECS), GitHub Actions, Terraform, CI/CD pipelines, and Helm/Argo
  • Real database experience with Postgres and DocumentDB (Mongo) across RDS/Aurora
  • A track record of working cross-functionally — managing timelines, setting expectations, and collaborating with other teams' engineers, not just your own
  • Already using AI as part of your actual workflow (coding, debugging, design)

Nice To Haves

  • Crossplane
  • Datadog

Responsibilities

  • Own reliability and observability across the organization — SLAs/SLOs, instrumentation, and on-call health
  • Design, deploy, and maintain Kubernetes infrastructure (Helm, EKS/ECS) and core AWS services (RDS/Aurora Postgres, networking, scaling)
  • Build and maintain CI/CD pipelines in GitHub Actions and Argo/Helm
  • Drive adoption of Infrastructure as Code standards (Terraform, and increasingly Crossplane) across multiple teams
  • Partner with 16+ engineering teams to roll out new standards, tools, and processes — adapting the rollout to how each team actually works
  • Participate in on-call rotation and lead incident response/root-cause analysis for issues that cross team boundaries
  • Use AI tools daily to work faster and more effectively, and help less AI-fluent teammates pick up the same habits
  • Represent SRE's perspective in planning conversations with engineering leadership

Benefits

  • Health, dental, and vision care for you and your family
  • Life insurance
  • Mental wellness coverage
  • Fertility and growing family support
  • Flex Time Off in addition to company-paid holidays
  • Paid family leave, medical leave, and bereavement leave policies
  • Retirement saving plans
  • Allowance to customize your work and technology setup at home
  • Annual professional development stipend
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service