Senior Systems Engineer, Diagnostics and System Behavior Analysis

AtomsSan Francisco, CA
$176,000 - $242,000Onsite

About The Position

Atoms is seeking a Senior Systems Engineer to own diagnostics and failure analysis across their development fleet. This role involves taking a failure from a log to an assigned root cause, managing triage classification, the severity model, and the diagnostic tooling. The goal is to build diagnostic coverage and automated classification to streamline failure analysis into a process that runs efficiently, ensuring only new failures require human intervention. The position spans multiple vehicle platforms with diverse hardware, compute, and software, and requires hands-on time in the garage and around vehicles.

Requirements

  • Bachelor's or master's degree in Computer Science, Computer Engineering, Electrical Engineering, Robotics, or a related field.
  • 6+ years in systems integration, diagnostics, or hardware and software interface engineering on vehicles.
  • L4 program or production ADAS experience.
  • A track record of diagnosing failures that spanned hardware and software, with specific examples where you established the root cause.
  • Ability to correlate hardware symptoms and video evidence with telemetry and bus data into a concrete timeline.
  • Sensor-level troubleshooting depth on cameras, lidar, radar, GNSS and IMU, including timing failure modes.
  • Strong working knowledge of vehicle networks and diagnostics, including CAN, CAN FD, Automotive Ethernet, and DTCs, and how failures present at each layer.
  • Python and command line log analysis at the level of building production tools.
  • Experience building automated triage or anomaly detection over fleet data.
  • Experience with monitoring and observability platforms and with log management at fleet scale.
  • Working experience with cloud storage and databases for telemetry: object storage, SQL, and time series or NoSQL stores.
  • Experience with a data collection or logging fleet, particularly diagnosing data path problems at volume.
  • Ability to work onsite five days a week in San Francisco, with periodic travel to test sites.

Responsibilities

  • Classify failures as software, firmware, hardware, data path, or system-level regressions, and route them with evidence.
  • Investigate novel failures using logs, telemetry, bus traces, and video, and establish the root cause.
  • Build parsing, correlation, and visualization tools to make large log datasets manageable for engineers outside the immediate role.
  • Develop frameworks to automatically detect known failure signatures, ensuring recurrence is caught by software.
  • Own monitoring for the logging path, including dropped frames, timestamp drift, bandwidth and storage limits, and alerting for vehicles producing unusable data.
  • Categorize severity consistently and identify failure clusters that should guide engineering priorities.
  • Set the standard for complete root cause analysis documentation and author troubleshooting guides.

Benefits

  • Medical, Dental, Vision, Disability, and Life Insurance
  • Flexible Spending Account / Health Savings Account Options
  • 401(k)
  • Equity
  • Sick Time, Unlimited Flexible Time Off, and Paid Holidays
  • Paid Parental Leave
  • Pre-Tax Commuter Benefit Plan
  • Team lunch in our SoMa office every Tuesday and Thursday
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service