Systems Design Engineer

TekWissenAustin, TX
Onsite

About The Position

The Data Center Platform Engineering Group (DPEG) is looking for an on-site Systems Design Engineer to work complex issues and problems as they arise in the lab or Data Center. The individual will be responsible for monitoring a ticketing system, debugging SW and HW problems, root causing, and fixing said issues. The role also involves monitoring the lab/data center for issues such as power outages, network outages, and liquid cooling leaks. The engineer must be able to physically install and replace hardware, and debug FW and OS issues, especially related to GPUs. This role requires the individual to act as both a technician and an engineer, with knowledge of Data Center systems that have CPUs and GPUs.

Requirements

  • Self-driven expertise in debugging problems and root causing issues.
  • Ability to quickly understand and apply specific details to tickets.
  • Strong communication skills and ability to work well as part of a team.
  • Proven track record of working in a Data Center on complex systems.
  • Knowledge in JIRA, scripting, Debug methodologies (BMC, BIOS, GPU).
  • Linux skills with the ability to run given automated scripts, look at logs, and give updates.
  • Experience working and communicating with OEM partners and vendors.

Nice To Haves

  • System level debug methodologies.
  • Proficiency with all standard hardware lab equipment (dediprog, scopes, forklifts, lift tools).
  • Strong analytical and problem-solving skills.
  • Proven ability to lead teams and drive issues to resolution.
  • Excellent verbal and written communication skills.

Responsibilities

  • Work issues through a JIRA ticket.
  • Debug platforms/systems and work the ticketing system to ensure priority items are getting attention.
  • Communicate clearly in tickets to explain the current state of an issue or systems.
  • Take direction from leaders based on priority needs of the business.
  • Work with partners, vendors, ODMs, OEMs on issues and ensure clear communication.
  • Drive continuous improvement within the organization.
  • Monitor the lab/data-center for any issues such as power outages, network outages, liquid cooling leaks, etc.
  • Physically install and replace hardware.
  • Debug FW and OS issues, especially related to GPUs.
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service