Vast.ai runs one of the largest GPU marketplaces in the world, with over 20,000 GPUs rented by more than 25,000 customers monthly for training, fine-tuning, and serving AI models. We are a profitable company with a flat team structure and rapid development cycles. This is an engineering role that operates in public, where you will be one of our heaviest users. A significant portion of your week will involve renting GPUs on Vast and building applications on them, such as serving open-source models with vLLM and SGLang, fine-tuning, running ComfyUI pipelines, and maintaining live inference endpoints. You will also be responsible for creating example repositories, Docker templates, and benchmarks for other developers. The remainder of your week will be dedicated to showcasing your work through short videos, technical write-ups, and engaging with the community on platforms like Discord and Reddit, as well as participating in hackathons. This role is not about building the Vast product itself, but rather about using it extensively in the open, identifying issues, and reporting back. If you are looking for a structured content calendar or a predefined campaign plan, this role may not be suitable. However, if you are someone who already experiments with GPUs on weekends to test new models, this role is for you.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Senior
Education Level
No Education Listed