Sr Staff DevOps Engineer

VantorWestminster, CO
$128,000 - $187,000Hybrid

About The Position

Vantor is seeking a Senior Staff DevOps Engineer with Site Reliability Engineering experience to join their Platform DevOps team. This role will support the build, deployment, reliability, and operations of the Vantor Hub software suite, which is part of Vantor’s Platform Infrastructure organization. The team works on a unified platform powering applications used by Vantor’s customers and is looking for a candidate inspired by working with world-class engineers to scale development, deployment, and operations efforts. The position is hybrid, requiring three days a week on-site in Westminster, CO. The Platform DevOps team designs, develops, secures, and operates custom and third-party software solutions in AWS, overseeing numerous AWS accounts across commercial and government environments. They guide DevOps, cloud infrastructure, and reliability engineering best practices, collaborating with hundreds of Vantor engineers to rapidly and securely release software changes into various AWS environments. This work is critical to the organization's velocity, enabling rapid introduction of business value without compromising quality or security. As a Senior Staff DevOps Engineer for Site Reliability, the individual will be instrumental in enhancing the reliability, scalability, observability, and operational maturity of Vantor Hub and its supporting infrastructure. This involves hands-on work with AWS, Kubernetes, infrastructure-as-code, CI/CD systems, monitoring platforms, and incident response processes, as well as influencing engineering standards and reliability practices across teams. This role is designed for an experienced engineer capable of independent operation, leading complex technical initiatives, mentoring others, and establishing best practices for production readiness, service reliability, automation, and cloud operations.

Requirements

  • Bachelor's Degree in Software Engineering, Computer Science, a related engineering field, or equivalent experience.
  • 8+ years of experience in DevOps, Platform Engineering, Site Reliability Engineering, or cloud operations, with demonstrated ownership of production systems.
  • Demonstrated experience owning, operating, and improving production systems in cloud-based environments.
  • Strong experience with Site Reliability Engineering practices, including service-level indicators, service-level objectives, error budgets, incident response, post-incident reviews, reliability metrics, and reliability-focused automation.
  • Strong proficiency with Amazon Web Services, including experience operating production workloads in multi-account AWS environments.
  • Strong proficiency with containerization technologies such as Docker and Kubernetes.
  • Strong proficiency with infrastructure-as-code (IaC) tools like Terraform or CloudFormation.
  • Proficiency in scripting languages such as Python, Bash, or PowerShell.
  • Solid understanding of networking concepts and protocols.
  • The ability to communicate and collaborate with team members and other colleagues.
  • Must be a U.S. citizen and be willing and able to obtain a U.S. Government security clearance.

Nice To Haves

  • Experience with high-availability systems, database replication, backup and restore, disaster recovery, and business-continuity planning.
  • Experience with zero-downtime or low-downtime deployment strategies, including blue/green deployments, canary releases, rolling deployments, feature flags, and automated rollback.
  • Experience with observability and incident management platforms such as Prometheus, Grafana, CloudWatch, ELK, PagerDuty, or similar tools.
  • Experience with secure cloud operations, least-privilege IAM, secrets management, vulnerability remediation, and audit-ready infrastructure.
  • Experience supporting government, regulated, or compliance-driven environments.
  • Practical experience using AI-assisted development or automation tools to improve engineering workflows, with a clear understanding of validation, review, privacy, and security considerations.
  • Active U.S. Government security clearance

Responsibilities

  • Lead reliability engineering efforts for the Vantor Hub platform and related infrastructure services.
  • Design, implement, and maintain scalable CI/CD pipelines for software and infrastructure delivery.
  • Build, operate, and improve cloud infrastructure automation using tools such as Terraform, CloudFormation, Kubernetes, Docker, and AWS-native services.
  • Troubleshoot complex infrastructure, deployment, networking, performance, reliability, and production issues.
  • Improve service availability, scalability, maintainability, reliability, security, and overall operational readiness.
  • Design and operate highly available, resilient, observable, and secure infrastructure across commercial and government AWS environments.
  • Partner with engineering teams to improve deployment safety, rollback capabilities, observability, production readiness, and operational support.
  • Participate in a team on-call rotation, currently approximately one week every twelve weeks, supporting production reliability and incident response.
  • Leverage AI development tools as force multipliers for software design, implementation, testing, and documentation while maintaining reliability, security, and engineering quality.
  • Contribute to shared engineering standards for responsible AI-assisted development, including validation practices, documentation expectations, and review patterns.
  • Use AI-assisted engineering tools to accelerate infrastructure automation, documentation, runbook development, test scaffolding, incident analysis, and repeatable operational workflows; validate outputs through peer review, testing, and security practices.

Benefits

  • Robust 401(k) with company match
  • Mental health resources
  • Student loan repayment assistance
  • Adoption reimbursement
  • Pet insurance
  • Incentive eligible with a target based on contribution, company performance, and/or individual results achieved
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service