About The Position

NVIDIA is seeking a Senior Backend/Platform Engineer to build and maintain the core infrastructure behind NVIDIA Brev. This role involves developing reliable cloud services, control planes, and execution environments that enable developers to access accelerated computing infrastructure across clouds. It's a high-impact position for an engineer who enjoys solving complex infrastructure problems, operating production systems, and building platforms for other engineers.

Requirements

  • B.S. degree or equivalent experience
  • 8+ years of relevant software engineering experience, with flexibility for exceptional candidates
  • Strong professional experience developing production systems in Go
  • Linux systems knowledge and the ability to debug across system layers
  • Strong networking fundamentals, including TCP/IP, DNS, routing, proxies, VPNs, and load balancing
  • Hands-on experience with Kubernetes and containerized workloads
  • Experience building infrastructure on AWS, GCP, or Azure
  • Backend or platform engineering experience with production systems
  • Strong distributed-systems fundamentals, including consistency, fault tolerance, concurrency, and failure handling
  • Experience building infrastructure, developer platforms, cloud services, or shared systems that other engineers depend on
  • A track record of owning reliability and operational outcomes in addition to feature delivery

Nice To Haves

  • Experience with Temporal or another durable workflow orchestration system
  • Experience designing multi-tenant platforms, control planes, or schedulers
  • Knowledge of VM lifecycle management, remote execution environments, or sandbox and isolation technologies
  • Experience with observability, reliability engineering, capacity planning, or infrastructure automation
  • Experience building AI agent platforms or developer execution environments
  • Experience building and operating GPU infrastructure, including GPU provisioning, scheduling, orchestration, or workload management

Responsibilities

  • Design, build, and operate production backend services and infrastructure in Go
  • Develop platform capabilities for provisioning, managing, and executing workloads across cloud environments
  • Build reliable control planes, APIs, schedulers, and infrastructure automation
  • Work deeply with Linux, Kubernetes, containers, networking, and public cloud infrastructure
  • Own systems throughout their lifecycle, including architecture, implementation, deployment, observability, incident response, and continuous improvement
  • Solve distributed-systems challenges involving state, concurrency, multi-tenancy, workload isolation, failure recovery, and scalability
  • Build infrastructure and platform primitives used by other engineers and developer-facing products
  • Establish best practices for system design, code quality, testing, reliability, and production operations
  • Collaborate across engineering and product teams to translate complex infrastructure requirements into simple, dependable developer experiences

Benefits

  • equity
  • benefits
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service