Technical Leader, Site Reliability Engineer

CiscoMilpitas, MI
Hybrid

About The Position

The application window is expected to close on: 08/24/2026. Job posting may be removed earlier if the position is filled or if a sufficient number of applications are received. Meet the Team: Cisco’s Collaboration Business Unit empowers people and organizations worldwide to connect, communicate, and innovate seamlessly. You will collaborate with a global team of software engineers and SREs responsible for delivering extraordinary collaboration experiences at scale. Our team supports backend services deployed worldwide and works closely with development, product, and operations partners to ensure reliability and performance. Webex is powering the shift to the hybrid workforce, helping people stay connected in a rapidly evolving digital world. We cultivate a startup-like culture that values innovation, ownership, and collaboration, while offering the scale and impact of a global technology leader. Who we are: Cisco Collaboration is the global engine of hybrid work, delivering the secure, AI-driven communication platform that connects millions of users and empowers the world’s largest enterprises to innovate together. The Persistence Team is a foundational component of this ecosystem. We provide Database-as-a-Service (DBaaS) at a large scale, managing distributed data footprints globally across private datacenters and AWS. We are responsible for the deployment, reliable operations, and resilience of the technology that powers the entire Cisco Collaboration Suite. Who You’ll Work With: You will join a globally distributed team of experts with diverse backgrounds who take pride in maintaining high availability with zero-downtime migrations. Our footprint is vast; you will collaborate with technical leads and architects across the entire Webex portfolio. You’ll also partner with our security teams to ensure our US Federal environments remain compliant. You’ll participate in a follow-the-sun on-call rotation during your regions daytime hours. About you: You are a distributed big data enthusiast who believes that if you have to do it twice, you should automate it. You thrive on solving complex "stateful" problems in a "stateless" world and are passionate about building resilient infrastructure that can survive regional outages.

Requirements

  • 8+ years hands on experience (deploy, monitor, scale, and upgrades) with distributed database technology such as CassandraDB, Opensearch, Kafka, PostgreSQL
  • 5+ Years managing AWS infrastructure at scale.
  • 5+ Years experience with infrastructure automation and DevOps practices such as CI/CD tools (Jenkins, GitHub Actions, GitLab CI) and Infrastructure as Code tools (Terraform, Ansible, Kubernetes), Python
  • 5+ years experience supporting live production systems for customer facing software as a service
  • 3+ years of experience leading technical projects, including responsibility for task estimation, milestone tracking, and stakeholder communication processes.

Nice To Haves

  • Exposure to modern data platoforms including data pipelines, streaming systems, AL/ML and GenAI workloads
  • Experience writing data driven applications in Java or Python and familiarity with data access patterns in software
  • Experience in regulated US Federal environments

Responsibilities

  • Design and implement large-scale solutions for Cassandra, Kafka, and OpenSearch across hybrid environments (AWS & Private Cloud).
  • Lead cloud engineering projects using Terraform, Ansible, and Kubernetes to improve service reliability and deployment velocity.
  • Lead capacity planning and performance tuning for high-throughput production systems.
  • Manage automated workflows for backups, security patching, and incident management with a focus on quality assurance in live environments.
  • Act as a consultant to application teams, guiding them on data access patterns and cloud migration strategies.

Benefits

  • medical, dental and vision insurance
  • a 401(k) plan with a Cisco matching contribution
  • paid parental leave
  • short and long-term disability coverage
  • basic life insurance
  • grants of Cisco restricted stock units
  • 10 paid holidays per full calendar year, plus 1 floating holiday for non-exempt employees
  • 1 paid day off for employee’s birthday, paid year-end holiday shutdown, and 4 paid days off for personal wellness determined by Cisco
  • 16 days of paid vacation time per full calendar year, accrued at rate of 4.92 hours per pay period for full-time employees (non-exempt)
  • flexible vacation time off program (exempt)
  • 80 hours of sick time off provided on hire date and each January 1st thereafter
  • up to 80 hours of unused sick time carried forward from one calendar year to the next
  • Additional paid time away may be requested to deal with critical or emergency issues for family members
  • Optional 10 paid days per full calendar year to volunteer
  • annual bonuses (non-sales roles)
  • performance-based incentive pay (sales roles)
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service