Manager of Platform DevOps

ZoomSan Jose, CA
$124,000 - $271,200Hybrid

About The Position

You will be the leader of our Platforms DevOps team. This team is responsible for Zoom’s Kubernetes clusters in both datacenters and clouds, as well as for our cloudops infrastructure. The scope and mandate of the team are broad, and you will have significant opportunities to partner with teams across Engineering, DevOps/SRE, and Security. You'll drive projects that enable new features, improve infrastructure reliability and security, and reduce costs. Within the team, you'll manage technical ICs and continuously improve processes with automation, policy, and AI. Broadly speaking, you are an exemplary SRE/DevOps leader with expertise running k8s, and knowledge of best practices for IaC, automation, monitoring, and security. You also have experience communicating nuanced concepts to diverse audiences (e.g. both engineers and senior management). You can balance technical tradeoffs, and can adapt plans to deliver quickly on the most important goals for our customers and partners. We manage Zoom's Kubernetes and cloud operations infrastructure at scale. Our team collaborates across Engineering and Security to automate processes and optimize platform performance. We enable reliable, secure, and cost-effective infrastructure for Zoom's global services.

Requirements

  • Hold BS/MS in Computer Science, related field, or equivalent practical experience
  • Have 5+ years of SRE or DevOps experience managing production infrastructure
  • Lead and develop technical teams with 3+ years of management experience
  • Have experience with at least one programming language, in addition to scripting languages
  • Have experience with cloud providers (e.g. AWS, OCI) and cloud infrastructure technologies (e.g. Terraform, kubernetes)
  • Demonstrate experience managing Kubernetes from an SRE perspective in production environments
  • Have experience with best practices for Continuous Deployment, including CI/CD tools and version control systems (e.g. Git, Jenkins, Argo CD, JFrog artifactory)
  • Be able to guide releases/deployments and/or participate in on-call shifts/incident response after hours or on weekends

Nice To Haves

  • Have experience with logging and monitoring tools (e.g. ELK stack, Prometheus, Grafana)
  • Have experience with system design and distributed computing at scale
  • Have experience with AI to automate SRE and management tasks
  • Have experience operating k8s in on-prem/colo datacenters
  • Ability to speak Chinese/Mandarin

Responsibilities

  • Leading projects that automate infrastructure management and advance the team's Kubernetes and cloud capabilities
  • Identifying and executing cost optimization opportunities across cloud resources and infrastructure
  • Driving security improvements across cloud and Kubernetes assets in partnership with Security teams
  • Collaborating with partner teams and stakeholders to understand needs and build project roadmaps
  • Developing team members and optimizing work processes through mature software practices and continuous improvement

Benefits

  • bonus
  • equity value
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service