SRE / Backend Engineer

NablaNew York, NY
$160,000 - $220,000Remote

About The Position

Nabla is seeking an SRE / Backend Engineer to join their Platform squad. This role involves owning production reliability, scaling the on-call culture, instrumenting the observability stack, driving the SRE roadmap, improving developer experience, and collaborating across engineering teams. The company is at an ambitious stage, backed by a $70M Series C, and is focused on building the next generation of clinical AI to improve healthcare. The Platform squad specifically handles the infrastructure and engineering foundations for Nabla's AI ambient documentation product, which serves over 100,000 clinicians and processes billions of tokens monthly.

Requirements

  • Demonstrates strong, senior-level experience in SRE, infrastructure, or back-end engineering, ideally within fast-paced, production-critical environments.
  • Hands-on experience with cloud infrastructure (GCP preferred; AWS or Azure experience considered if you're ready to learn a new stack)
  • You've been part of a structured on-call process: alerting policies, runbooks, incident post-mortem, escalation flows
  • Strong software engineering fundamentals - you write clean, maintainable infrastructure-as-code and are as comfortable reviewing a Terraform PR as you are debugging a live incident
  • Autonomy is your default mode - you can identify what needs to be done, prioritize it yourself, and drive it to completion without being managed step by step
  • You communicate clearly and effectively with engineers and leadership across distributed teams, including during incidents, even outside regular working hours
  • Candidates must be currently authorized to work in the United States and must be able to maintain authorization to work in the United States without requiring employer sponsorship, now or in the future. For this role, the Company is not able to sponsor, transfer, or extend employment-based visas or other work authorization.

Nice To Haves

  • experience with Kubernetes
  • GCP-native tooling (Cloud Run, GKE, Pub/Sub)
  • PostgreSQL at scale
  • security/compliance in a healthcare context

Responsibilities

  • Own production reliability - take meaningful accountability for the stability, robustness, and performance of our platform across the full stack
  • Own and scale the on-call culture - take what's already in place and bring it to the next level: strengthen our runbooks, sharpen escalation paths, and build incident response processes that can scale with the team
  • Instrument everything - develop and improve our monitoring, alerting, and observability stack so that issues surface before they reach users, and our engineers can diagnose them fast
  • Drive the SRE roadmap - work with the Platform squad and Engineering leadership to define and prioritize the reliability investments that matter most, from SLOs to chaos engineering
  • Make engineers faster - reduce toil and improve the developer experience: better tooling, clearer feedback loops, smoother deployments
  • Collaborate across squads - partner with back-end, ML, and front-end engineers to embed reliability best practices into how we build, not just how we operate

Benefits

  • Competitive salary and stock options
  • 100% individual coverage for Medical, Dental, and Vision insurance
  • Unlimited paid time off
  • 11 national holidays
  • Unlimited sick leave
  • Paid leave for new parents
  • $1,500 to purchase home office equipment
  • Full ownership of your time and schedule
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service