Diagnostic & System Engineer

Foxconn-PCE TechnologySanta Clara, CA
$130,000 - $155,000Onsite

About The Position

The Diagnostic & System Engineer is a factory-based engineer who modifies and validates the portion of test and diagnostic programs used on the AI server production and repair floor. Rather than owning full test-program development, this role focuses on real-time modification, tuning, and validation of factory test scripts, diagnostic routines, test sequences, and parameters to keep production running, reproduce failures, and improve fault coverage. In addition to test-program work, the role performs system-level diagnostics and failure isolation across hardware, firmware, software, network, storage, GPU, thermal, and power domains. The AI Server Diagnostic & System Engineer works closely with Test Engineering, Repair, Manufacturing, Quality, Engineering, Program Management, suppliers, and customer-facing teams to improve diagnostic accuracy, reduce no-trouble-found conditions, and accelerate failure resolution.

Requirements

  • Bachelor's degree in Electrical Engineering, Computer Engineering, Manufacturing Engineering, or a related technical field required.
  • Associate degree with relevant test, diagnostics, repair, or manufacturing engineering experience may be considered.
  • 4-7 years of experience in system diagnostics, server debug, test engineering, failure analysis, repair engineering, or related technical fields.
  • Demonstrated experience troubleshooting complex hardware, firmware, BIOS, BMC, Linux, networking, storage, GPU, power, thermal, or system-level failures using logs and diagnostic tools.
  • Hands-on experience modifying and validating test programs or diagnostic scripts in a production or lab environment; AI server, rack-level system, and manufacturing test-flow experience preferred.

Nice To Haves

  • AI server, rack-level system, and manufacturing test-flow experience preferred.

Responsibilities

  • Modify, validate, and release factory test programs, diagnostic scripts, test sequences, and parameters under defined change control to support AI server production, test, and repair operations.
  • Reproduce failures, collect logs, analyze symptoms and error codes, and isolate hardware, firmware, software, network, storage, GPU, thermal, and power issues.
  • Validate diagnostic changes through retest, correlation, and verification before production release, addressing test escapes, false failures, coverage gaps, and complex failures outside standard repair processes.
  • Develop and maintain diagnostic procedures, troubleshooting trees, debug guides, failure-isolation methods, test coverage, change logs, and technical escalation records.
  • Partner with Test Engineering, Repair, Manufacturing, Quality, Engineering, and suppliers on root cause, corrective action, NPI/ramp readiness, customer escalations, and technician training.
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service