Quince is seeking a Staff ML Engineer to join their growing team. The ideal candidate is a deeply technical ML infrastructure engineer who combines hands-on mastery with system-level thinking. This role involves building and operating production-grade ML systems at scale, including distributed training pipelines, feature stores, and high-throughput inference serving. The engineer will be responsible for engineering platforms that other engineers love to use, designing for extensibility, observability, and resilience. The role requires an individual who gravitates toward the hardest problems, such as optimizing GPU utilization, designing zero-downtime model deployment systems, and defining architectural patterns for industrializing AI at scale. The position operates with high autonomy, demands exceptional standards, and involves elevating other engineers through code reviews, technical mentorship, and setting a high bar for quality.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Senior
Education Level
No Education Listed