Engineering Manager, Cloud Infrastructure

Replit•Foster City, CA
•$250,000 - $325,000•Hybrid

About The Position

Replit is seeking a hands-on Engineering Manager to lead their Cloud Infrastructure team. This role involves overseeing shared infrastructure as code (IaC), networking, storage, compute, and service mesh platforms that are critical for Replit's product and platform teams. The manager will lead and grow an existing engineering team responsible for building and operating these foundational elements, including Kubernetes, shared edge networking, service mesh, and workload identity. This is a platform-building role with production accountability, requiring the ability to delve into designs and incidents while fostering a team that can operate independently.

Requirements

  • Demonstrated engineering management experience, including leading and developing engineers, making prioritization and performance decisions, hiring thoughtfully, and delivering through a team.
  • Software-oriented infrastructure depth, with experience building and operating cloud platforms or distributed systems.
  • Ability to reason across infrastructure code, Kubernetes, networking, service identity, and stateful dependencies.
  • Safe-change and production judgment, including experience owning consequential migrations and incidents.
  • Ability to explain failure modes and rollback limits.
  • Understanding of when simplifying a system is better than adding another platform.
  • Platform-product and engineering judgment, including understanding internal customers, creating interfaces other teams adopt, and making clear tradeoffs among reliability, developer autonomy, engineering effort, and workload efficiency.

Nice To Haves

  • Experience with multi-tenant, cellular, regional, or dedicated enterprise infrastructure.
  • Familiarity with GCP/GKE, Terraform or similar IaC systems, Cloudflare, Envoy/Istio, SPIFFE/SPIRE, and managed data services.
  • Experience with large-fleet rightsizing, infrastructure consolidation, or migrating CI compute without disrupting developer workflows.
  • A track record using AI tools to increase engineering output while preserving production safeguards.

Responsibilities

  • Own the cloud-platform roadmap, leading the team's IaC, networking, storage, compute, and service mesh platforms.
  • Translate product, platform, reliability, and security needs into sequenced outcomes, balancing foundational investment, lifecycle work, and delivery commitments against team capacity.
  • Make infrastructure repeatable and self-service by building maintained IaC interfaces for services, cells, connectivity, identities, and shared resources.
  • Enable internal customer teams to provision infrastructure without bespoke coordination or dependence on individual experts.
  • Operate what the team builds, owning platform availability, upgrades, isolation, recovery, and incident remediation.
  • Maintain clear SLOs, sustainable on-call coverage, and primary and backup owners for critical systems.
  • Stay technically engaged by reviewing designs and production changes, debugging difficult failure modes, and using AI coding tools to prototype and automate.
  • Apply rigorous review and verification to AI-generated infrastructure changes.
  • Build and grow a high-ownership engineering team by coaching engineers, developing technical leaders, setting clear expectations, managing performance, and hiring.
  • Delegate meaningful ownership as the team grows.

Benefits

  • Competitive Salary & Equity
  • 401(k) Program with a 4% match (US Only)
  • Health, Dental, Vision and Life Insurance
  • Short Term and Long Term Disability
  • Paid Parental, Medical, Caregiver Leave
  • Flexible Time Off (FTO) + Holidays
  • Commuter Benefits (In-Office & US Only)
  • Monthly Wellness Stipend
  • Autonomous Work Environment
  • In Office Set-Up Reimbursement (In-Office Only)
  • Quarterly Team Gatherings
  • In Office Amenities (In-Office Only)
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service