Principal Site Reliability Engineer

UnitedHealth Group•Eden Prairie, MN
•$134,600 - $230,800

About The Position

Optum Tech is a global leader in health care innovation. Our teams develop cutting-edge solutions that help people live healthier lives and help make the health system work better for everyone. From advanced data analytics and AI to cybersecurity, we use innovative approaches to solve some of health care’s most complex challenges. Your contributions here have the potential to change lives. Ready to build the next breakthrough? Join us in making healthcare work better for everyone through people-led, responsible AI while Caring. Connecting. Growing together. We are seeking a highly experienced Principal Site Reliability Engineering (SRE) leader to drive reliability, resilience, and secure software engineering practices across all critical applications within the Digital Consumer Engineering organization using applied AI. In this role, you will operate at an enterprise and portfolio level, shaping reliability strategy, influencing architectural decisions, and enabling engineering teams to build and operate highly resilient, secure, and scalable platforms. You will play a critical role in advancing operational excellence, strengthening platform resilience, and supporting long-term business growth.

Requirements

  • Undergraduate degree in applicable area of expertise or equivalent experience
  • 8+ years of relevant experience in SRE, software engineering, or infrastructure engineering
  • 5+ years of experience operating at enterprise scale influencing strategy across multiple teams
  • 4+ years of experience with in distributed systems, cloud infrastructure, and observability
  • Proven track record of driving large-scale reliability or security transformation
  • Demonstrated solid leadership and cross-functional influence skills

Nice To Haves

  • Advanced degree
  • Experience leading enterprise transformation initiatives
  • Experience with AI/ML applied to operations or security
  • Vendor/platform strategy experience

Responsibilities

  • Define and drive enterprise-wide SRE strategy, standards, and operating models across digital consumer platforms
  • Establish and champion a culture of reliability engineering, security by design, and operational excellence
  • Influence architecture, platform design, and investment decisions to improve reliability, scalability, and resilience at scale
  • Partner with senior leaders across engineering, security, and operations to align reliability priorities with business outcomes
  • Identify and incubate AI-driven capabilities to advance reliability, observability, automation, and proactive risk management
  • Provide technical leadership and governance for reliability practices across systems including SLO frameworks and availability targets
  • Drive enterprise-wide adoption of automation and self-healing systems
  • Lead proactive risk management including threat modeling and resilience testing
  • Establish reliability metrics and reporting frameworks to drive measurable improvements
  • Partner with enterprise security leadership to embed security engineering practices into platform design
  • Drive proactive threat detection and response strategies leveraging automation and AI
  • Influence and standardize secure development and operational practices
  • Lead cross-functional initiatives to strengthen preventative security posture
  • Partner with platform teams to define scalable and resilient infrastructure strategies
  • Drive adoption of infrastructure-as-code and automation practices
  • Establish operational maturity and continuous improvement frameworks

Benefits

  • comprehensive benefits package
  • incentive and recognition programs
  • equity stock purchase
  • 401k contribution
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service