Senior Infrastructure Engineer

LevelSeattle, WA

About The Position

As an Senior Infrastructure Engineer on the Platform team, you will architect, build, and operate the cloud infrastructure and developer platform that every Level product runs on. You will own critical infrastructure end-to-end — the Kubernetes platform, infrastructure-as-code, CI/CD and GitOps delivery, networking (ingress and egress), DNS, observability, and cloud security posture — and provide the reliable, self-service foundations the rest of engineering builds on. You will work on a small, senior-leaning team where infrastructure decisions have direct, visible impact on reliability, performance, cost, and developer velocity.

Requirements

  • 5+ years operating large-scale cloud infrastructure (AWS strongly preferred)
  • Deep IaC experience (Terraform/OpenTofu; CloudFormation/Pulumi/CDK also relevant)
  • Strong Docker/Kubernetes (EKS) production experience
  • Scripting/automation proficiency (Python, Go, or Bash)
  • Solid cloud networking fundamentals (VPC, DNS, load balancing, ingress, firewalls/WAF, VPNs) and security best practices
  • Proven CI/CD and GitOps experience (GitHub Actions or similar)
  • Observability experience (metrics/logs/traces) used to drive real decisions
  • Track record leading infrastructure projects independently, end to end
  • Strong communication across technical and non-technical audiences

Nice To Haves

  • ArgoCD, Atlantis, Linkerd/Envoy, Traefik, Karpenter, Helm
  • OpenTelemetry, SigNoz (our stack), Grafana, Datadog, or Prometheus
  • Backstage or other internal developer platform experience
  • AI/ML infra experience (GPU scheduling, model/agent hosting, inference gateways)
  • Rust service CI/CD, CloudFront/CDN experience
  • AWS Solutions Architect / DevOps Engineer – Professional certification
  • Distributed-systems background, OSS infrastructure contributions

Responsibilities

  • Design, build, and operate secure, highly available AWS infrastructure using Terraform/OpenTofu with a GitOps workflow (Atlantis).
  • Own capacity planning, DR, and cost optimization for the systems you run.
  • Operate and evolve EKS: autoscaling (Karpenter), upgrades, core add-ons, and Helm-based delivery (ArgoCD).
  • Build and maintain GitHub Actions pipelines that let platform and product teams ship fast and safely, with self-service tooling where it makes sense.
  • Own ingress/egress (Traefik), service mesh and mTLS (Linkerd/Envoy), load balancing, edge TLS, and DNS (Route 53, Terraform-managed).
  • Build observability with OpenTelemetry and SigNoz; use telemetry to drive reliability, performance, and cost decisions.
  • Serve as an escalation point for complex incidents, leading troubleshooting and post-mortems.
  • Apply cloud security best practices across identity, secrets, and network boundaries, with particular care for student data and K-12 privacy.
  • Operate posture/vulnerability tooling (Security Hub, GuardDuty, Inspector, Snyk) and org guardrails (Control Tower, SCPs).
  • Set standards, mentor engineers, and leave the platform better than you found it.
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service