Senior Staff DevOps Engineer

SailPointAustin, TX
Remote

About The Position

As a Senior Staff DevOps Engineer on the Infrastructure Platform team, you will be a technical leader responsible for designing, building, and operating SailPoint's global Identity Security Cloud infrastructure on AWS. You will partner with engineering teams across the US, India, EMEA, and APAC to deliver a resilient, secure platform. You will serve as the Kubernetes platform leader for the team, driving large-scale improvements such as implementing a service mesh (Istio, Linkerd, or AWS App Mesh) across hundreds of microservices and production EKS clusters, and guiding engineers on cloud-native deployment patterns. You will also play a key role in supporting SailPoint's PCI compliance initiative, helping ensure our platform infrastructure meets PCI DSS requirements. This is a fully remote position for candidates based in the USA or Canada. The role carries significant technical influence across the organization, setting architecture direction, mentoring engineers, and driving operational excellence without requiring people management. The ideal candidate is a self-starter who thrives in complex, fast-paced SaaS environments, has proven hands-on experience rolling out and operating a service mesh at scale across large microservices environments, brings expert-level Kubernetes and AWS cloud knowledge, values quality and reliability, and is excited to work on an infrastructure team pushing the boundaries of cloud-native operations.

Requirements

  • Strong interpersonal and teaming skills, with the ability to set and enforce process and influence engineers across teams and geographies.
  • Ability to operate effectively in an agile, entrepreneurial environment with global stakeholders.
  • Prior experience as a technical lead or Staff+ IC in a global engineering organization.
  • 3+ years of hands-on experience designing, implementing, and operating a service mesh at scale in production Kubernetes environments.
  • 10+ years of experience in 24x7 production operations, supporting highly available SaaS or cloud service environments.
  • 10+ years of experience with containerization, virtualization, and configuration management technologies.
  • 5+ years of hands-on experience with Kubernetes in production at scale.
  • 5+ years of experience with Terraform (IaC), managing infrastructure across multiple AWS accounts and regions.
  • 5+ years of experience designing and implementing CI/CD pipelines, especially for Terraform, Kubernetes, and microservices.
  • 5+ years of experience with scripting/programming languages (Python, Go, or similar) and strong shell scripting proficiency.
  • Strong understanding of Linux, networking, distributed systems, and production troubleshooting.
  • Demonstrated experience scaling a service mesh across a large microservices fleet, including phased adoption strategy, sidecar resource management, control plane scaling, and performance tuning under high request volume.
  • Experience with monitoring and logging stacks (e.g., Prometheus, Grafana, OpenSearch or equivalent).

Nice To Haves

  • Experience with multi-cluster or multi-region service mesh federation (e.g., Istio multi-cluster, Linkerd multicluster extension) in production.
  • Familiarity with compliance frameworks in regulated enterprise SaaS, including PCI DSS and FedRAMP-adjacent practices, with experience implementing or operating platforms subject to PCI compliance requirements preferred.

Responsibilities

  • Own and lead the full lifecycle of service mesh implementation at scale, from architecture and rollout planning through production operations across hundreds of microservices running on EKS.
  • Establish service mesh standards and governance adopted organization-wide, including onboarding runbooks, sidecar injection policies, traffic policy templates, and documented failure-mode playbooks for production incidents.
  • Mentor and upskill engineers across teams on service mesh architecture, troubleshooting at scale, performance tuning, and capacity planning for mesh-heavy environments.
  • Design, operate, and optimize production Kubernetes clusters at scale on AWS (EKS), including cluster architecture, upgrades, node management, networking, storage, and multi-tenant isolation patterns.
  • Define and drive Kubernetes standards and best practices across teams, including workload design, resource management, security hardening (RBAC, Pod Security Standards, network policies), Helm/chart conventions, and deployment patterns.
  • Design and scale infrastructure to meet rapidly increasing customer demand, data sovereignty requirements, and regional expansion.
  • Automate deployment, monitoring, incident response, and capacity management using GitOps and CI/CD best practices.
  • Develop and improve operational practices, runbooks, and platform engineering standards.
  • Collaborate with development teams to bring new features and services into production safely and efficiently.
  • Proactively meet information security and compliance standards (e.g., PCI DSS), including supporting SailPoint's PCI compliance initiative through secure platform design, controls implementation, and audit readiness.
  • Participate in and help improve the on-call rotation; drive post-incident reviews and systemic fixes.

Benefits

  • Health and wellness coverage: Medical, dental, and vision insurance
  • Disability coverage: Short-term and long-term disability
  • Life protection: Life insurance and Accidental Death & Dismemberment (AD&D)
  • Additional life coverage options: Supplemental life insurance for employees, spouses, and children
  • Flexible spending accounts for health care, and dependent care; limited purpose flexible spending account
  • Financial security: 401(k) Savings and Investment Plan with company matching
  • Time off benefits: Flexible vacation policy
  • Holidays: 8 paid holidays annually
  • Sick leave
  • Parental support: Paid parental leave
  • Employee Assistance Program (EAP) and Care Counselors
  • Voluntary benefits: Legal Assistance, Critical Illness, Accident, Hospital Indemnity and Pet Insurance options
  • Health Savings Account (HSA) with employer contribution
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service