The On-Device Machine Learning team at Apple is responsible for enabling the Research to Production lifecycle of cutting-edge machine learning models that power magical user experiences on Apple's hardware and software platforms. The team builds critical infrastructure that begins with onboarding the latest machine learning architectures to Apple devices, optimization toolkits to optimize these models to better suit the target devices, machine learning compilers and runtimes to execute these models as efficiently as possible, and the benchmarking, analysis and debugging toolchain needed to improve on new model iterations. This infrastructure underpins most of Apple's critical machine learning workflows across Camera, Siri, Health, Vision, etc., and as such is an integral part of Apple Intelligence. Our group is seeking an Engineering Manager to lead the Performance Tools and Services team, with a focus on the tools, services, and infrastructure that make on-device ML performance measurable, understandable, and improvable. The team is responsible for the frontend web services for introspecting ML models and their on-device execution, the backend web services that power them, and the on-device toolchain that gathers low-level performance data, associates it with high-level (PyTorch) framework-level ops, and reports it. The team is also responsible for the infrastructure for running ML inference across fleets of devices. We are building the first end-to-end developer experience for ML development that, by taking advantage of Apple's vertical integration, allows developers to iterate on model authoring, optimization, transformation, execution, debugging, profiling and analysis. This role focuses on giving ML developers fast, accurate, and actionable insight into how their models execute on Apple devices. We're looking for a manager that has proven experience in and passion for providing high quality developer tools and capabilities in the fast paced and dynamic space of ML. As the manager in this role, you will lead a diverse team spanning full-stack web development, backend services, distributed systems, and low-level on-device performance tooling. You will partner with leaders across the organization and company to develop our platform while supporting clients internally and externally. The role requires a solid technical understanding of ML execution on device, performance analysis and profiling, and the systems that connect low-level signals to framework-level (e.g., PyTorch) semantics.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Manager
Education Level
Ph.D. or professional degree