AI Software Engineer, Lightspeed Studios

NVIDIA•Santa Clara, CA
•$224,000 - $431,250

About The Position

At NVIDIA Lightspeed Studios, we are passionate about pushing the limits of technology and are looking for a Senior Software Engineer to join our team. We help shape the future of games by combining the best AI and graphics technology NVIDIA has to offer with the most advanced games and tools in the industry. In this role, you'll help take new, unannounced projects powered by state-of-the-art AI models from research to real-time, interactive gaming experiences. If you're passionate about bringing generative AI to games, we'd love to hear from you!

Requirements

  • BS, MS, or PhD in Computer Science, Electrical Engineering, or a related field, or equivalent experience.
  • 12+ years of software engineering experience, including substantial hands-on work shipping AI, deep learning, or machine learning systems.
  • Expert-level C++ and Python.
  • Hands-on experience training and fine-tuning deep learning models with PyTorch or a similar framework, including distributed training.
  • Experience in one or more of these areas: video generation, world models, diffusion or flow-matching models, transformers, or neural rendering.
  • Experience optimizing GPU inference with CUDA, TensorRT, TensorRT-LLM, or Triton.
  • A track record of turning complex prototypes into shipped products, and clear communication across research, engineering, and art.

Nice To Haves

  • Hands-on experience with interactive, action-conditioned, or real-time world models or video generation.
  • Experience developing custom CUDA or Triton kernels.
  • Experience with game engine development or real-time rendering pipelines.
  • Open-source contributions (such as Diffusers, FastVideo, vLLM, SGLang, or TensorRT-LLM), publications, or patents.

Responsibilities

  • Build and ship software for upcoming, not-yet-announced projects that bring world models and video diffusion models to real-time gaming, from prototype to release.
  • Train, fine-tune, and evaluate models, including data curation and adapting models to game-specific content and controls.
  • Optimize models and inference for latency, token throughput, and quality, using techniques such as distillation, quantization, and reduced-step sampling on RTX devices and in the cloud.
  • Profile and remove bottlenecks across the stack, from model architecture to GPU kernels.
  • Integrate models into Unreal Engine, Unity, and custom engines, so they run efficiently alongside rendering.
  • Lead technical decisions, mentor other engineers, and collaborate with NVIDIA Research, rendering, and art teams.

Benefits

  • equity
  • benefits
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service