Principal Site Reliability Engineer

MiniMed•Los Angeles, CA
•$138,000 - $261,000

About The Position

As a Principal Site Reliability Engineer (SRE), you will own the end-to-end reliability, scale, and performance of our global healthcare platform built on Amazon Web Services (AWS). This role bridges the gap between software engineering and systems operations, with a hyper-focus on maintaining HIPAA compliance, HITRUST standards, and cybersecurity frameworks. You will architecture automated, self-healing platforms that ensure our life-critical applications remain highly available 24x7.

Requirements

  • Requires a Bachelor's Degree and minimum of 7 years of relevant experience, or advanced degree with a minimum of 5 years of relevant experience.

Nice To Haves

  • Deep architectural knowledge of core AWS services, including networking (VPC), computing (EC2, ECS), IAM permission isolation, and data layers (RDS, DynamoDB).
  • 5+ years of dedicated experience in Site Reliability Engineering or DevOps supporting production workloads at scale.
  • Direct experience operating platforms within a regulated environment subject to HIPAA, SOC2, or FDA medical device software regulations (IEC 62304).
  • Heavy production experience configuring and troubleshooting containerized workloads via Amazon ECS.
  • Comfortable writing robust systems tooling or automation scripts in Python, Go, or Bash.

Responsibilities

  • Design, scale, and maintain highly available, fault-tolerant infrastructure on AWS. Ensure all underlying architecture aligns with necessary security protocols.
  • Eliminate manual operational tasks. Champion GitOps practices by treating infrastructure exclusively as code using tools like CloudFormation.
  • Build and mature deep telemetry frameworks using NewRelic, CloudWatch and OpenSearch to actively trace application health and catch system regressions before they hit patient care workflows.
  • Lead high-severity incident responses as an Incident Manager. Lead Blameless Postmortems and Root Cause Analysis (RCA) to drive long-term engineering remediations.
  • Ensure technical safeguards are enforced across AWS data stores containing Protected Health Information (PHI). Support compliance engineering teams by generating data encryption evidence and access logs for regulatory audits.
  • Co-design cloud-native microservices alongside our Core Engineering teams, building automated CI/CD guardrails that allow developers to deploy code safely and autonomously.

Benefits

  • health, dental, and vision insurance
  • Health Savings Account
  • Healthcare Flexible Spending Account
  • life insurance
  • long-term disability leave
  • dependent daycare spending account
  • incentive plans
  • 401(k) plan with company match
  • short-term disability coverage
  • paid time off and holidays
  • Employee Stock Purchase Plan
  • Employee Assistance Program
  • Non-qualified Retirement Plan Supplement
  • Capital Accumulation Plan
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service