Rivian's Autonomy org needs a Staff Software Engineer, Compute & Storage to own the distributed compute and storage platform that every autonomy workload runs on. This sits in the Platform Services team in the AI Platform organization in the Autonomy team. Autonomy training, simulation base evaluation / validation, and autonomy visualization all depend on the same two things: available compute and fast access to data. The role requires deep expertise in Kubernetes-based distributed compute, large-scale object storage, and the performance and cost tradeoffs of running both at petabyte scale. You'll work with the AI Platform, Perception, Planning, Simulation, and Vehicle Integration, Product Management, and other technology partners to operate a platform serving a fleet of over 100,000 vehicles, hundreds of petabytes of drive and simulation data, and training clusters of thousands of GPUs across multiple clouds. This is a platform ownership role, and it's measured by what it lets other engineers do: how fast someone goes from idea to trained model or data pipeline, how many scenarios simulation runs per day, and what each of those costs.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Senior
Education Level
Associate degree