The On-Device Machine Learning team at Apple transforms groundbreaking research into practical applications, enabling billions of Apple devices to run powerful AI models locally, privately, and efficiently. This team operates at the intersection of research, software engineering, hardware engineering, and product development. They build essential infrastructure for machine learning at scale on Apple devices, including onboarding innovative architectures to embedded systems, developing optimization toolkits for model compression and acceleration, building ML compilers and runtimes for efficient execution, and creating comprehensive benchmarking and debugging toolchains. This infrastructure supports Apple’s machine learning workflows across Camera, Siri, Health, Vision, and other core experiences, contributing to the Apple Intelligence ecosystem. The role is for an ML Infrastructure Engineer with a focus on model compilation, working closely with model authoring, runtime, and performance teams to ensure models can leverage the full capabilities of the hardware. The team is building an end-to-end developer experience for machine learning development using Apple’s vertical integration, covering model authoring, optimization, transformation, execution, debugging, profiling, and analysis. This specific role focuses on the core runtime for execution across various devices and use cases. The ideal candidate is a creative, versatile, and passionate software engineer interested in machine learning, common compiler optimizations, and system software engineering. The team utilizes an MLIR-based compiler stack to target the neural engine, GPU, and CPU for ML workflows and execution.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Mid Level
Education Level
No Education Listed