Cloud Operations Engineer

Florida BlueUnited States,
$109,300 - $177,600Hybrid

About The Position

We're moving workloads to AWS, and this is the person who keeps them running afterward: patched, backed up, monitored, and recoverable. The Cloud Operations Engineer owns that ongoing operational work and builds automation so it doesn't keep growing headcount as the estate grows. This role is a standard business-hours on eastern standard time zone engineering role with a shared on-call rotation, not a NOC or shift job. The focus is building automation and operational standards, not watching dashboards.

Requirements

  • 6+ years related work experience in infrastructure operations or systems engineering with 4+ years experience operating cloud workloads
  • 1+ years direct supervisory/management experience
  • Related Bachelor's degree required
  • Strong hands-on AWS experience running production workloads.
  • Proficiency in a programming or scripting language such as Python or Go.
  • Hands-on Terraform and CI/CD pipeline experience.
  • Experience defining and operating against SLOs and error budgets.
  • Observability tooling experience with CloudWatch, Prometheus, Grafana, Datadog, or equivalent.
  • Incident management and on-call experience in a production environment.
  • Kubernetes or EKS in production.

Nice To Haves

  • Multi-region or disaster recovery design, including failover testing.
  • Healthcare, financial services, or other regulated industry experience.
  • AWS GovCloud experience. Chaos engineering or resilience testing.
  • AWS certification such as DevOps Engineer, SysOps Administrator, or Solutions Architect.

Responsibilities

  • Build and maintain monitoring and alerting for migrated workloads. Alert on conditions that require action.
  • Automate patching, backup validation, certificate renewal, and other recurring operational tasks.
  • Define and enforce operational readiness criteria a workload must meet before it goes live in cloud.
  • Build and test disaster recovery runbooks. Validate recovery time and recovery point targets through failover exercises.
  • Respond to incidents, participate in the on-call rotation, and drive root cause to closure.
  • Convert manual procedures into executable automation and self-healing where the task is repeatable.
  • Track availability and operational health against defined targets. Report on recurring issues.
  • Manage backup, retention, and restore testing for cloud workloads.
  • Troubleshoot and maintain AWS compute, including Windows Server and Linux instances.
  • Partner with infrastructure operations to transition migrated workloads into steady-state support.
  • Maintain runbooks and operational documentation a new on-call engineer can use unaided.

Benefits

  • Medical, dental, vision, life and global travel health insurance
  • Income protection benefits: life insurance, short- and long-term disability programs
  • Leave programs to support personal circumstances
  • Retirement Savings Plan including employer match
  • Paid time off, volunteer time off, 10 holidays and 2 well-being days
  • Additional voluntary benefits available; and a comprehensive wellness program
  • Competitive pay
  • Opportunities for incentive or commission compensation
  • Regular annual reviews with pay for performance considerations for base pay increases
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service