Sr Platform DevOps Engr, SRE - Remote

UnitedHealth Group•Eden Prairie, MN
•$91,700 - $163,700•Remote

About The Position

Optum Tech is a global leader in health care innovation. Our teams develop cutting-edge solutions that help people live healthier lives and help make the health system work better for everyone. From advanced data analytics and AI to cybersecurity, we use innovative approaches to solve some of health care’s most complex challenges. Your contributions here have the potential to change lives. Ready to build the next breakthrough? Join us to start Caring. Connecting. Growing together. We are seeking a Senior Platform DevOps Engineer with Site Reliability Engineering (SRE) experience to help build, automate, secure, and operate enterprise cloud platforms and services. This role will be responsible for designing and maintaining cloud infrastructure, deployment automation, observability solutions, security controls, operational processes, and platform standards that enable development teams to deliver secure, reliable, and scalable applications. The ideal candidate is a hands-on engineer who thrives in complex cloud environments and has experience supporting production platforms, modern CI/CD practices, infrastructure as code, monitoring and alerting, cloud security, incident response, and operational excellence initiatives. You’ll enjoy the flexibility to work remotely from anywhere within the U.S. as you take on some tough challenges. For all hires in the Minneapolis or Washington, D.C. area, you will be required to work in the office a minimum of four days per week.

Requirements

  • 5+ years of experience in DevOps, Site Reliability Engineering (SRE), Platform Engineering, Cloud Engineering, or related technology roles
  • 4+ years of hands-on experience supporting production workloads in Microsoft Azure
  • 4+ years of experience implementing Infrastructure as Code using Terraform
  • 4+ years of experience designing and supporting CI/CD pipelines using GitHub Actions
  • 1+ years of experience with container technologies such as Docker and Kubernetes
  • 1+ years of experience with cloud networking, identity management, secrets management, and platform security
  • 1+ years of experience implementing monitoring, logging, alerting, dashboards, and operational observability solutions
  • 1+ years of experience supporting production incidents, troubleshooting complex technical issues, and driving root cause analysis
  • 1+ years of experience remediating security vulnerabilities and implementing secure software delivery practices
  • 1+ years of experience with Git-based source control, branching strategies, repository governance, and release management
  • 1+ years of experience with scripting and automation using PowerShell, Python, Bash, or similar languages

Nice To Haves

  • Bachelor's degree in Computer Science, Engineering, Information Technology, or related field, or equivalent practical experience.
  • Experience with Kubernetes platform administration and GitOps practices
  • Experience with cloud-native monitoring and observability platforms
  • Experience supporting highly available, mission-critical production systems
  • Experience with application security scanning, dependency management, and software supply chain security
  • Experience implementing SLOs, SLIs, error budgets, and other reliability engineering practices
  • Experience with disaster recovery planning, resiliency testing, and capacity management
  • Experience working in regulated enterprise environments
  • Experience with AI-assisted development, platform automation, or operational tooling
  • Prior experience serving as a technical lead, mentor, or senior engineering resource within a platform engineering, DevOps, or SRE organization
  • Proven understanding of cloud infrastructure, reliability engineering, security, and operational excellence principles

Responsibilities

  • Design, build, and maintain cloud infrastructure using Infrastructure as Code (IaC) practices
  • Develop, enhance, and support CI/CD pipelines and deployment automation across multiple environments
  • Implement platform standards, reusable automation frameworks, and engineering best practices
  • Support cloud-native application deployments and containerized workloads
  • Implement and maintain monitoring, logging, alerting, and observability solutions to improve system reliability and operational visibility
  • Define and support SRE practices including service health monitoring, operational readiness, incident response, runbooks, and reliability improvements
  • Troubleshoot and resolve complex application, infrastructure, networking, identity, security, and deployment issues
  • Manage and improve platform security by implementing security controls, vulnerability remediation processes, secret management, and automated security scans
  • Partner with development, security, architecture, and operations teams to improve platform reliability, scalability, and delivery speed
  • Support disaster recovery, resiliency, backup, recovery, and business continuity initiatives
  • Participate in on-call and incident response activities as required
  • Automate repetitive operational tasks using scripting and cloud automation tools
  • Drive continuous improvement across DevOps, SRE, platform engineering, cloud operations, and governance practices
  • Mentor engineers and promote DevOps and SRE best practices across teams
  • Evaluate and leverage AI-assisted tools and automation to improve engineering productivity, operational efficiency, and software delivery processes

Benefits

  • comprehensive benefits package
  • incentive and recognition programs
  • equity stock purchase
  • 401k contribution
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service