Machine Learning Researcher, Foundation Models [SWE Org]

Cupertino Until 9/22/2026 First posted July 24, 2026 Last posted July 24, 2026
Job description

We build frontier foundation models that power intelligent experiences at Apple. Our team works across the full training lifecycle: including pre-training foundation models, and developing mid-training approaches that bridge general capability and task-specific performance. What makes our work distinct is that we're engineering models specifically for Apple silicon and optimized for experiences that are private, personal, and deeply integrated into the OS. We're solving frontier problems in reward modeling to resist reward hacking, handling sparse and delayed rewards in agentic settings, and aligning models reliably across the spectrum from open-ended creative tasks to precise, action-taking workflows. If you're drawn to hard problems where the research and the product are inseparable, this is the team.

Description

We believe that the most interesting problems in deep learning research arise when we try to apply learning to real-world use cases, and this is also where the most important breakthroughs come from. You will work with a close-knit and fast growing team of world-class engineers and scientists to tackle some of the most challenging problems in foundation models and deep learning.

Further, you will have opportunities to identify and develop novel applications of deep learning in Apple products. You will see your ideas improve the experience of billions of users.

Minimum Qualifications

Demonstrated expertise in deep learning with publication record in relevant conferences (e.g., NeurIPS, ICML, ICLR, COLM, ACL, NAACL, EMNLP, ACL) or a track record in applying deep learning techniques to products
Proficient programming skills in Python and one of the deep learning toolkits such as JAX, PyTorch, or Tensorflow
Ability to work in a collaborative environment.
Code large language models.
PhD, or equivalent practical experience, in Computer Science, or related technical field.

Preferred Qualifications

Reinforcement learning, on-policy distillation.
Post-training, mid-training large language models.
LLM context lengthening.

About this role

Summary

Develop and optimize foundation models with deep learning, reinforcement learning, and language modeling.

Job title

Machine Learning Researcher, Foundation Models [SWE Org]

Experience level

doctorate or equivalent practical experience

Industry

software

Location requirements

Cupertino; on-site work only

Salary

Not specified

Management role

No

Skills & keywords

Required skills

deep learningPythonJAXPyTorchTensorFlowlarge language models

Preferred skills

reinforcement learningon-policy distillationmodel trainingcontext lengthening

Specializations

deep learninglarge language modelsreinforcement learningcontext lengthening
Locations

Structured locations inferred from the posting.

Cupertino, CA, USA

On-site City