Sr Staff DevOps Engineer

VantorWestminster, CO
Hybrid

About The Position

Vantor is forging the new frontier of spatial intelligence, helping decision makers and operators navigate what’s happening now and shape what’s coming next. Vantor is a place for problem solvers, changemakers, and go-getters—where people are working together to help our customers see the world differently, and in doing so, be seen differently. Come be part of a mission, not just a job, where you can: Shape your own future, build the next big thing, and change the world. To be eligible for this position, you must be a U.S. Citizen. This position requires an active U.S. Government security clearance, applicants who do not currently hold the required clearance will not be eligible for consideration. Employment for cleared roles is contingent upon verification of clearance status. Export Control/ITAR: Certain roles may be subject to U.S. export control laws, requiring U.S. person status as defined by 8 U.S.C. 1324b(a)(3). Please review the job details below. Vantor’s Platform DevOps team is growing and we’re looking for a Senior Staff DevOps Engineer with Site Reliability Engineering experience to support the build, deployment, reliability, and operations of the Vantor Hub software suite. This team is part of Vantor’s Platform Infrastructure organization, working on a unified platform powering applications used by Vantor’s customers. We’re looking for a candidate that is inspired by working with a talented group of world-class engineers, scaling our development, deployment, and operations efforts. This position is hybrid with three days a week on-site with your colleagues in Westminster, CO. The Platform DevOps team designs, develops, secures, and operates a wide range of custom and third-party software solutions deployed in AWS. The team oversees dozens of AWS accounts between commercial and government environments, and guides DevOps, cloud infrastructure, and reliability engineering best practices. We collaborate with hundreds of Vantor engineers to quickly and securely build, test, and release software changes into many commercial and government AWS environments. This software and adjacent services are critical to the overall velocity of the organization and allows our teams to introduce business value rapidly without sacrificing quality or security. As a Senior Staff DevOps Engineer for Site Reliability, you will play a key role in improving the reliability, scalability, observability, and operational maturity of Vantor Hub and its supporting infrastructure. You will work hands-on with AWS, Kubernetes, infrastructure -as- code, CI/CD systems, monitoring platforms, and incident response processes while also influencing engineering standards and reliability practices across teams. This role is intended for an experienced engineer who can operate independently, lead complex technical initiatives, mentor others, and help establish best practices for production readiness, service reliability, automation, and cloud operations.

Requirements

  • Bachelor's Degree in Software Engineering, Computer Science, a related engineering field, or equivalent experience.
  • 8 + years of experience in DevOps, Platform Engineering, Site Reliability Engineering, or cloud operations, with demonstrated ownership of production systems.
  • Demonstrated experience owning, operating, and improving production systems in cloud-based environments.
  • Strong experience with Site Reliability Engineering practices, including service-level indicators, service-level objectives, error budgets, incident response, post-incident reviews, reliability metrics, and reliability-focused automation.
  • Strong proficiency with Amazon Web Services, including experience operating production workloads in multi-account AWS environments.
  • Strong proficiency with containerization technologies such as Docker and Kubernetes.
  • Strong proficiency with infrastructure-as-code (IaC) tools like Terraform or CloudFormation.
  • Proficiency in scripting languages such as Python, Bash, or PowerShell.
  • Solid understanding of networking concepts and protocols.
  • The ability to communicate and collaborate with team members and other colleagues.
  • Must be a U.S. citizen and be willing and able to obtain a U.S. Government security clearance.

Nice To Haves

  • Experience with high-availability systems, database replication, backup and restore, disaster recovery, and business-continuity planning.
  • Experience with zero-downtime or low-downtime deployment strategies, including blue/green deployments, canary releases, rolling deployments, feature flags, and automated rollback.
  • Experience with observability and incident management platforms such as Prometheus, Grafana, CloudWatch, ELK, PagerDuty, or similar tools.
  • Experience with secure cloud operations, least-privilege IAM, secrets management, vulnerability remediation, and audit-ready infrastructure.
  • Experience supporting government, regulated, or compliance-driven environments.
  • Practical experience using AI-assisted development or automation tools to improve engineering workflows, with a clear understanding of validation, review, privacy, and security considerations.
  • Active U.S. Government security clearance.

Responsibilities

  • Lead reliability engineering efforts for the Vantor Hub platform and related infrastructure services.
  • Design, implement, and maintain scalable CI/CD pipelines for software and infrastructure delivery.
  • Build, operate, and improve cloud infrastructure automation using tools such as Terraform, CloudFormation, Kubernetes, Docker, and AWS-native services.
  • Troubleshoot complex infrastructure, deployment, networking, performance, reliability, and production issues.
  • Improve service availability, scalability, maintainability, reliability, security, and overall operational readiness.
  • Design and operate highly available, resilient, observable, and secure infrastructure across commercial and government AWS environments.
  • Partner with engineering teams to improve deployment safety, rollback capabilities, observability, production readiness, and operational support.
  • Participate in a team on-call rotation, currently approximately one week every twelve weeks, supporting production reliability and incident response.
  • Leverage AI development tools as force multipliers for software design, implementation, testing, and documentation while maintaining reliability, security, and engineering quality.
  • Contribute to shared engineering standards for responsible AI-assisted development, including validation practices, documentation expectations, and review patterns.
  • Use AI-assisted engineering tools to accelerate infrastructure automation, documentation, runbook development, test scaffolding, incident analysis, and repeatable operational workflows; validate outputs through peer review, testing, and security practices.

Benefits

  • robust 401(k) with company match
  • mental health resources
  • student loan repayment assistance
  • adoption reimbursement
  • pet insurance
  • incentive eligible with a target based on contribution, company performance, and/or individual results achieved
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service