At SF Tensor, we're building the future of high-performance compute by rethinking and rebuilding the AI stack from the hardware to the cloud. Our Kernel Optimizer finds the fastest possible form for code on any vendor and cluster topology, and our Model Foundry manages runs, simplifies research, and moves workloads across clouds and chips. We are backed by prominent investors and are looking for individuals who believe in the necessity of compute advancements for AI progress. This role focuses on the modeling side of our enterprise offering, which promises to turn a customer's dataset into a specialist model in days. We achieve this through techniques like SFT, RL, DPO, and distillation. The infrastructure is exceptionally strong, enabling rapid model training. Experiments are managed through Model Foundry, providing a versioned and reproducible environment for running experiments on optimal hardware. This powerful engine has been used for post-training large language models, robotics models, and pre-training AlphaFold v3.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Senior
Education Level
No Education Listed