Member of Technical Staff

WaferSan Francisco, CA
Onsite

About The Position

Member of Technical Staff Inference performance depends on how efficiently models use the underlying hardware. Techniques across kernels, compilers, runtimes, and serving systems can dramatically improve latency and throughput, but these optimizations are difficult, hardware-specific, and slow to reproduce across new accelerators. And none of it counts until it is running in a customer's production traffic. At Wafer, we are building AI systems that automatically optimize inference workloads across silicon. The goal is fungible token capacity. Any accelerator optimized toward serving inference most efficiently. Wafer is well funded and serves trillions of tokens a month for mission critical workloads. We serve the highest performance inference to fast-growing AI startups. Members of Technical Staff build the systems that make that possible and own the customers running on them. There is no separate solutions team, and no layer between you and the workload.

Requirements

  • Exceptional engineer
  • Shipped and operated production systems
  • Can open a profiler, read a trace, and find the problem yourself
  • Can be handed an unfamiliar system and own it inside a week
  • Owned a customer relationship
  • Credible in a room full of engineers
  • Can take a skeptical ML team through a benchmark, defend the methodology, and concede the point when they are right
  • Work without a spec. The problem arrives half-defined from a customer who does not yet know what they need, and you come back with a scoped answer rather than a list of questions

Nice To Haves

  • Inference experience helps

Responsibilities

  • Develop and optimize high-performance computing kernels, and work across inference engine internals and serving infrastructure
  • Develop AI agents to do autonomous inference engineering
  • Design, deploy, and operate heterogeneous clusters across vendors
  • Own customer accounts end to end. Sales gets the first meeting. From there you decide what to prove, you build it, you keep it running in production, and you carry the relationship
  • Win the technical evaluation. Prove Wafer on the customer's own workload rather than a synthetic benchmark, and be the person who can explain the result to their engineers
  • Own production for your accounts. When latency moves or error rates climb, you find it, you fix it or route it, and you are who the customer hears from
  • Inform product roadmap. You sit closer than anyone to how Wafer behaves under real load, and the engineering team builds against what you report

Benefits

  • $200-300K base salary + generous equity
  • Fully covered medical, dental, and vision insurance
  • Daily lunch and dinner
  • Unlimited PTO
  • Parental leave
  • $1K/month housing stipend (post-tax) if you live within walking distance (0.5 miles) from the office
  • Covered Uber/Waymo from/to office
  • Visa sponsorship available
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service