GPU Systems Engineer – HPC / Parallel Computing

Vast.aiSan Francisco, CA
Onsite

About The Position

Vast.ai is seeking a systems engineer with HPC or parallel programming experience to help scale AI inference. The role involves leveraging knowledge of high-performance systems to optimize GPU performance at the bleeding edge of AI. This is a full-time, on-site position at either their San Francisco or Los Angeles offices. The tech stack includes CUDA/C++, GPGPU, Python, and Linux.

Requirements

  • Advanced C++ (C++17/20 preferred)
  • Expertise with at least one parallel framework (CUDA, HIP, SYCL, OpenCL, OpenACC, or similar)
  • Strong background in systems optimization and HPC performance tooling

Nice To Haves

  • Familiarity with distributed training/inference frameworks

Responsibilities

  • Design and optimize GPU kernels and tensor libraries
  • Translate HPC techniques into scalable AI inference solutions
  • Evaluate emerging architectures and resource management approaches
  • Collaborate with technical leadership to improve GPU infrastructure efficiency

Benefits

  • Comprehensive health, dental, vision, and life insurance
  • 401(k) with company match
  • Meaningful early-stage equity
  • Onsite meals, snacks, and close collaboration with founders/tech leaders
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service