About The Position

Join Apple's Multimodal Intelligence team during an exciting era of Artificial Intelligence to help deliver groundbreaking Apple products and experiences. The team builds and ships the Computer Vision and Machine Learning systems behind Apple Intelligence, covering data collection and curation, training and fine-tuning, evaluation, optimization, and on-device deployment. They have a proven history of integrating Apple's sensing hardware with large foundation models to create features like Visual Intelligence and on-device foundation models for text and visual understanding across various Apple devices. The focus is on developing experiences where devices can perceive, comprehend, and reason about their surroundings privately, responsively, and on-device whenever feasible.

Requirements

  • Experience in deep learning with demonstrated work in at least one area of multimodal systems (e.g., vision, language, video, audio, etc.).
  • Proficiency in Python and in a modern deep learning framework such as PyTorch or JAX.
  • Experience with rapid prototyping, reproduction, and validation of research ideas.
  • Ability to work in a collaborative environment.
  • Ability to communicate the results of analyses in a clear and effective manner.
  • BS and a minimum of 3 years of relevant industry experience.

Nice To Haves

  • Master's or PhD, or equivalent practical experience, in Computer Science, Computer Vision, Machine Learning, or related technical field.
  • Deep expertise in multimodal foundation models, with a focus on practical applications.
  • Track record of translating research into practical applications either through published work or industry experience.
  • Strong applied research experience in at least one major area of model development (data curation, pre-training, fine-tuning, alignment, or evaluation), particularly as it applies to multimodal systems.
  • Experience with large-scale training pipelines, including working with large datasets and scaling models across distributed systems.
  • Experience bridging research ideas with production constraints.

Responsibilities

  • Build the pipelines, infrastructure, and production systems to transform multimodal foundation models into shipping Apple Intelligence features.
  • Own end-to-end model delivery, including building and scaling data curation and training pipelines.
  • Fine-tune and optimize large multimodal models for on-device and hybrid execution.
  • Establish reproducible evaluation and regression testing for text and visual understanding.
  • Harden promising approaches into robust, maintainable production systems considering real-world constraints like latency, memory, power, and privacy.
  • Collaborate with modeling, platform, hardware, and product engineering teams across Apple.
  • Incorporate future hardware design and product needs into implementation decisions.
  • Collaborate broadly to deliver the best possible products.
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service