Site Reliability Engineer

Empower
$87,400 - $123,400Hybrid

About The Position

Our vision for the future is based on the idea that transforming financial lives starts by giving our people the freedom to transform their own. We have a flexible work environment, and fluid career paths. We not only encourage but celebrate internal mobility. We also recognize the importance of purpose, well-being, and work-life balance. Within Empower and our communities, we work hard to create a welcoming and inclusive environment, and our associates dedicate thousands of hours to volunteering for causes that matter most to them. Chart your own path and grow your career while helping more customers achieve financial freedom. Empower Yourself. Applicants must be authorized to work for any employer in the U.S. We are unable to sponsor or take over sponsorship of an employment visa at this time, including CPT/OPT.

Requirements

  • Experience maintaining high availability and resiliency within AWS infrastructure components, including EKS, EC2, RDS, S3, VPC, and others
  • Proficiency with Infrastructure as Code frameworks such as Terraform and Cloudformation
  • Demonstrated experience with containerization and orchestration technologies such as Docker and Kubernetes
  • Experience with technologies, systems, networks, and potential gaps that can impact an organization’s ability to effectively detect and respond to production incidents
  • Strong problem-solving abilities and a desire to learn

Nice To Haves

  • Bachelor’s degree in Computer Science, Information Systems or equivalent experience
  • AWS, Kubernetes, or relevant certifications
  • Experience with observability suites and APM tooling such as DataDog, AppDynamics, New Relic, etc.
  • Strong programming skills in one or more languages such as shell, Go, Python, etc.
  • Experience supporting Java Spring Boot applications
  • Production experience in Kubernetes, especially EKS

Responsibilities

  • Establish key indicators (SLIs) measuring the performance of services and build proactive monitor and alerts
  • Support projects of varying complexity and impact across multiple disciplines and teams
  • Conduct in-depth analysis of problems to identify relevant findings and root causes
  • Play a key role in ensuring the high availability, resilience, and scalability of containerized applications in production
  • Document critical systems and create runbooks for incidents
  • Lead capacity planning and right-sizing exercises
  • Maintain and optimize infrastructure as code (IaC)
  • Troubleshoot and resolve complex system and deployment issues
  • Manage observability within Kubernetes, specifically EKS
  • Collaborate with development teams to support releases and create highly scalable, resilient, and maintainable services
  • Work in a GitOps driven environment

Benefits

  • Medical, dental, vision and life insurance
  • Retirement savings – 401(k) plan with generous company matching contributions (up to 6%), financial advisory services, potential company discretionary contribution, and a broad investment lineup
  • Tuition reimbursement up to $5,250/year
  • Business-casual environment that includes the option to wear jeans
  • Generous paid time off upon hire – including a paid time off program plus ten paid company holidays and three floating holidays each calendar year
  • Paid volunteer time — 16 hours per calendar year
  • Leave of absence programs – including paid parental leave, paid short- and long-term disability, and Family and Medical Leave (FMLA)
  • Business Resource Groups (BRGs) – BRGs facilitate inclusion and collaboration across our business internally and throughout the communities where we live, work and play. BRGs are open to all.
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service