About The Position

The Evaluation team builds and evolves the evaluation ecosystem that powers developing and scaling GM’s autonomous driving technology. We develop metrics, automated workflows, and analysis approaches that enable data-driven decisions across AV development and verification. Partnering with Autonomy, Simulation, Systems, and Safety teams, we act as system-level integrators and arbiters of end-to-end AV quality. We own large scale test scenario libraries, continuous evaluation pipelines, and critical risk assessment and release gating components, treating road testing, data mining, training, and metrics as first-class use cases in a unified analytics framework. By joining this team, you will help shape GM’s core evaluation platforms, turn system-level results into clear feedback, and help accelerate validated AV deployment at scale.

Requirements

  • 7 + years of applied experience develo ping complex evaluation, simulation, or test frameworks .
  • Proficient in developing Python for production systems, including unit testing, code review, performance tradeoffs, and reliability best practices.
  • Proven cross-team technical leadership, including defining strategies adopted by multiple teams and influencing system and architecture decisions.
  • Strong written and verbal communication , driving decisions, communicating risk, and giving constructive feedback to diverse stakeholders.
  • B achelor’s or higher degree in Computer Science, Engineering, or equivalent experience.

Nice To Haves

  • Experience in autonomous driving or high-stakes field robotics; designing, running, and interpreting large-scale simulation and field experiments.
  • Experience w orking on test strategies and validation for safety-critical products .
  • A strong, data-driven curiosity to investigate anomalies and systematically root-cause discrepancies.
  • Familiarity with SQL, time-series data analysis, performance monitoring tools and dashboarding systems (e.g., Looker, Streamlit ).

Responsibilities

  • Architect large- scale evaluation pipelines that quantify the accuracy and reliability of simulation tests used for autonomous vehicle software validation.
  • Lead cross-functional initiatives with Autonomy, Systems Engineering, Simulation, and Data teams to tightly integrate team-owned evaluation products into regular development workflows and release decision processes.
  • Invent novel methodologies and deliver implementation to quantify and characterize the trustworthiness and effectiveness of simulation testing and evaluation products at scale .
  • Drive technical roadmaps and strategic priorities while partnering cross-functionally to integrate new simulation technologies aligned with AV goals.
  • Own and refine key simulation evaluation metrics and KPIs used for readiness and safety decisions; synthesize and present results and tradeoffs to stakeholders; make insights readily available to partner teams through interactive dashboards.
  • Maintain a high technical standard through architectural design, design reviews, and code reviews, setting patterns and best practices for the broader team.

Benefits

  • medical
  • dental
  • vision
  • Health Savings Account
  • Flexible Spending Accounts
  • retirement savings plan
  • sickness and accident benefits
  • life insurance
  • paid vacation & holidays
  • tuition assistance programs
  • employee assistance program
  • GM vehicle discounts
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service