We're building the evaluation platform that will serve all of Apple's generative AI and agent systems. Evaluating non-deterministic AI systems is one of the hardest unsolved problems in production ML — and one Apple has to get right at scale. We're building the platform that makes it tractable for every team here. This is a hands-on engineering role with a lot of autonomy. You'll write a lot of Python and own meaningful pieces of the platform end-to-end. You'll be partnering closely with research engineers, model and serving teams, product and feature teams, and the infra and data platform groups this work integrates with.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Senior
Education Level
No Education Listed