Quince is seeking a Staff ML Engineer to join their growing team. The ideal candidate is a deeply technical ML infrastructure engineer who combines hands-on mastery with system-level thinking. This role involves building and operating production-grade ML systems at scale, including distributed training pipelines, feature stores, and high-throughput inference serving. The engineer will be responsible for engineering platforms that other engineers love to use, designing for extensibility, observability, and resilience. This role is for someone who gravitates toward the hardest problems, such as optimizing GPU utilization, designing zero-downtime model deployment systems, or defining architectural patterns for industrializing AI at scale. The position requires high autonomy, exceptional standards, and the ability to elevate engineers through code reviews, technical mentorship, and by setting a high bar for quality.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Senior
Education Level
No Education Listed