About The Position

Vantor is seeking a Senior Platform Engineer, AI Infrastructure to help build and operate the secure, scalable foundation for our AI-driven development organization. This role supports Site Sentry, Maritime Sentry, and Storyline—products delivered as part of the broader Tensorglobe spatial intelligence platform. You will work at the intersection of software engineering, cloud infrastructure, reliability, automation, and performance. You will help development teams move quickly and safely by improving the systems that build, deploy, monitor, and operate mission-critical applications across cloud, on-premises, and edge environments. The ideal candidate is a strong software engineer who can design and develop production-quality solutions, automate repetitive operational work, instrument systems for observability, troubleshoot complex failures, and continuously improve platform performance. Our codebase is primarily Python, and success in this role requires a pragmatic, hands-on approach to engineering and operations.

Requirements

  • Put security first: protect company, customer, mission, and operational data through secure-by-design engineering, least-privilege access, secrets management, threat-aware automation, and strong security and compliance practices.
  • Design, implement, and maintain reliable CI/CD pipelines for software projects, including automated testing, quality gates, artifact management, release promotion, rollback, and deployment verification.
  • Automate deployment, scaling, configuration, and monitoring of cloud-based infrastructure using technologies such as Kubernetes, Docker, Terraform, and Google Cloud services including Cloud Run.
  • Troubleshoot infrastructure, deployment, availability, latency, capacity, and performance issues across distributed systems.
  • Collaborate with software development teams to optimize code and services for performance, scalability, reliability, operability, and cost efficiency.
  • Implement and enforce best practices for security, compliance, change management, incident response, and operational readiness.
  • Continuously evaluate and recommend improvements to DevOps, SRE, developer productivity, observability, and platform engineering processes and tools.
  • Demonstrated experience as a full-stack or backend software developer, with the ability to write maintainable production code—not only configuration and infrastructure definitions.
  • Professional experience developing in Python and working with APIs, services, databases, automated tests, and source-control workflows.
  • Strong communication and collaboration skills, with the ability to explain technical tradeoffs and partner effectively with engineering, product, security, and customer-facing teams.

Nice To Haves

  • Experience with geospatial information systems (GIS), mapping, spatial data, or location-based applications.
  • Experience working with satellite imagery, aerial imagery, remote sensing, computer vision, or other earth-observation data.
  • Prior military, intelligence, defense, national security, or mission operations experience.
  • Experience operating software in disconnected, air-gapped, classified, or otherwise highly controlled environments.
  • Experience with AI/ML platform operations, model-serving infrastructure, evaluation pipelines, or responsible AI practices.
  • Experience with PostgreSQL or other production databases, query and schema optimization, caching, and data lifecycle management.
  • Experience with infrastructure-as-code, GitOps, policy-as-code, and supply-chain security.

Responsibilities

  • Own and improve platform reliability through SRE practices, including service-level objectives, monitoring, alerting, incident response, root-cause analysis, and operational reviews.
  • Lead performance management and optimization for services, infrastructure, Cloud Run workloads, databases, and supporting platform components.
  • Build and maintain deployment automation that makes releases repeatable, observable, secure, and easy to recover.
  • Develop internal tools and platform capabilities in Python to reduce toil, improve developer experience, and make the right operational behavior the easiest behavior.
  • Design and operate observability across logs, metrics, traces, health checks, dashboards, and actionable alerts.
  • Partner with developers to improve application architecture, code quality, resource utilization, scalability, and production readiness.
  • Build, deploy, and maintain on-premises environments at customer sites, including installation, upgrades, configuration, diagnostics, and lifecycle support.
  • Support deployment models that span cloud, customer data centers, and edge environments, adapting engineering practices to security, connectivity, and operational constraints.
  • Participate in technical design reviews, incident response, documentation, and continuous improvement initiatives across the platform and product lifecycle.

Benefits

  • robust 401(k) with company match
  • mental health resources
  • student loan repayment assistance
  • adoption reimbursement
  • pet insurance
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service