About The Position

As a Site Reliability Engineering (SRE) Technical Leader on the Intersight Team, you will play a key role in ensuring the reliability, scalability, and security of our cloud platforms. The broader team is composed of experienced engineers who value innovation and accountability. You will represent the Intersight SRE team, working in a dynamic environment, tackling challenges with creativity, providing technical leadership in defining and delivering on the team's technical roadmap. You will collaborate with cross-functional teams, including software development, product management, customers, and security teams, to design, influence, build, and maintain SaaS systems operating at multi-region scale. Your work will directly impact the success of our initiatives by ensuring the underlying platform infrastructure is robust, efficient, and aligned with operational excellence. We are seeking an experienced Engineer to encourage and represent a high-performing team dedicated to ensuring the reliability and scalability of cloud services, with a focus on a rapidly growing next-generation project. The ideal candidate will have hands-on SRE or systems/network administration experience, with familiarity in AWS. This role involves close collaboration across product engineering, service engineering, and SRE teams in a high-trust, well-coordinated environment.

Requirements

  • Bachelor’s +12 years of related experience, or Master’s +8 years of related experience or PhD +5 years of related experience
  • Experience designing and implementing scalable, reliable, and production-ready solutions.
  • Experience with cloud platforms, preferably AWS, and Infrastructure as Code (IaC) using Terraform or Ansible.
  • Experience with Linux, Docker, Kubernetes/EKS, networking, security, and CI/CD.
  • Experience with observability and monitoring tools
  • Proficiency in Python, Go, or similar programming languages, with a strong understanding of software development principles.

Nice To Haves

  • Experience building/managing a cloud-based data platform, automation and orchestration of their infrastructure and maintaining high availability, system reliability at scale.
  • Ability to handle multiple competing priorities in a fast-paced environment
  • Experience in architecting software and infrastructure at scale with a sense of ownership and accountability.
  • Strong passion for learning, researching, and developing innovative technologies that deliver significant customer impact.
  • Exceptional communication and presentation skills, with the ability to translate complex technical concepts for non-experts while fostering collaboration and driving excellence within diverse technical teams.
  • Experience in providing technical leadership and mentoring engineers.
  • Experience with cloud and infrastructure security.
  • Certifications: CKA (Certified Kubernetes Administrator), CKAD (Certified Kubernetes Application Developer), AWS Certified DevOps Engineer, or equivalent certifications in cloud and security domains.

Responsibilities

  • Design, build, and optimize cloud and data infrastructure to ensure the high availability, reliability, and scalability of systems to meet customer needs, while implementing SRE principles such as monitoring, alerting, error budgets, and fault analysis.
  • Collaborate closely with multi-functional teams, including customers, development, product management, and security teams, to create secure, scalable solutions and enhance operational efficiency through automation.
  • Troubleshoot complex technical problems in production environments, perform root cause analyses, and contribute to continuous improvement efforts through postmortem reviews and proactive performance optimization.
  • Lead the architectural vision and shape the team’s technical strategy and roadmap, balancing immediate needs with long-term goals, driving innovation, and influencing the technical direction.
  • Serve as a mentor and technical leader, guiding teams and fostering a culture of engineering and operational excellence by sharing your deep knowledge and experience.
  • Engage with customers and stakeholders to understand use cases and feedback, translating them into actionable insights and effectively influencing at all levels.
  • Apply your strong programming skills to integrate software and systems engineering, building core data platform capabilities and automation to meet enterprise customer needs and roadmap objectives.
  • Develop strategic roadmaps, processes, plans, and infrastructure to efficiently deploy new software components at an enterprise scale while enforcing engineering standards.

Benefits

  • medical, dental and vision insurance
  • a 401(k) plan with a Cisco matching contribution
  • paid parental leave
  • short and long-term disability coverage
  • basic life insurance
  • 10 paid holidays per full calendar year, plus 1 floating holiday for non-exempt employees
  • 1 paid day off for employee’s birthday, paid year-end holiday shutdown, and 4 paid days off for personal wellness determined by Cisco
  • 16 days of paid vacation time per full calendar year, accrued at rate of 4.92 hours per pay period for full-time employees (non-exempt employees)
  • flexible vacation time off program (exempt employees)
  • 80 hours of sick time off provided on hire date and each January 1st thereafter, and up to 80 hours of unused sick time carried forward from one calendar year to the next
  • Additional paid time away may be requested to deal with critical or emergency issues for family members
  • Optional 10 paid days per full calendar year to volunteer
  • annual bonuses
  • performance-based incentive pay
  • Cisco restricted stock units
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service