Apple's AIML Evaluation team builds the systems and methodologies that measure and improve the quality of foundation models and agentic experiences. We are looking for a senior, hands-on Machine Learning Engineering Manager to lead a small team working at the intersection of model evaluation, agent optimization, and data generation. In this role, you will help define how evaluation closes the loop with model and product development, turning observed quality gaps into targeted improvements to prompts, agent harnesses, datasets, and models. You will combine technical depth with people leadership. You should be comfortable moving from research papers and experimental results to production-quality ML pipelines, while mentoring engineers and aligning teams around a clear technical direction. Your work will span Apple Foundation Models and product teams, with the goal of creating repeatable evaluation and refinement loops that improve the quality of Apple intelligence experiences.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Senior