Machine Learning Scientist - Apple Services Engineering, GenAI & ML Frameworks
Apple New York, New York, United States
Computers and Electronics Manufacturing · 10,001+ employees
About the role
The role involves bridging foundation model capabilities with production systems through LLM training, agentic system optimization, and deployment-aware engineering. You will lead cross-functional initiatives to integrate cutting-edge models into user-facing features while ensuring production readiness and scalability.
What they look for
Requirements
Candidates must possess a BS, MS, or PhD in a quantitative field and demonstrate proficiency in Python and deep learning frameworks like PyTorch, Jax, or Tensorflow. A proven track record in training large-scale models or building distributed systems is essential for this position.
Full description
Apple Services GenAI & ML Frameworks team aims at bridging foundation model capabilities with real-world production systems. The work spans LLM continual pretraining, posttraining, agentic reinforcement learning, agentic system optimization etc.. This role is part of the cross-LOB effort to support various GenAI use cases across ASE, and specializes in improving LLM domain knowledge, tool use, reasoning, and system integration—working closely with product, infra, and foundation model teams to bring cutting-edge models into user-facing features at scale.
Description
We are seeking a strong candidate who can operate end-to-end across model development and production integration—someone equally strong in (1) LLM training (domain-adaptive continual pretraining, post-training, preference optimization / RL such as GRPO-style methods), (2) agentic systems (tool schemas, multi-turn reliability, rubric- or verifier-based learning loops), and (3) deployment-aware optimization (latency/cost/reliability tradeoffs, evaluation harnesses, and iterative improvement from production signals).
The ideal candidate has a track record of turning LLM research into shipped capabilities, can partner effectively with product, infra, and foundation model teams, and can lead ambiguous cross-LOB initiatives from problem definition through execution and scaling. Experience building robust tooling around synthetic data generation, eval, and training pipelines for LLMs is strongly preferred, since this role is expected to raise the bar on both research velocity and production readiness.
Minimum Qualifications
BS/MS in a quantitative field, including Computer Science, Maths, Statistics, Physics, etc. Proficient programming skills in Python Hands-on experience working with deep learning toolkits such as Jax, Tensorflow or PyTorch Proven track record in training or deployment of large models or building large-scale distributed systems Deep understanding of Deep Learning and Large Language Models (LLMs) Natural Language Processing
Preferred Qualifications
PhD in a quantitative field, including Computer Science, Maths, Statistics, Physics, etc.
Similar roles
-
Machine Learning Co-Op (Fall 2027)
Hendrickson Canton, Ohio, United States
-
Machine Learning Engineer (5-8 yrs)
Advanced Space Westminster, Colorado, United States · $124K–$171K/yr
-
Senior Machine Learning Engineer - New Verticals Agentic Foundations
DoorDash USA San Francisco, California, United States · $137K–$299K/yr
-
Senior Machine Learning Engineer, Payments
Airbnb United States · $191K–$223K/yr
-
Artificial Intelligence and Machine Learning Engineer, Senior
Booz Allen Hamilton Wahiawa, Hawaii, United States · $99K–$225K/yr
-
Machine Learning Engineer
Booz Allen Hamilton Aurora, Colorado, United States · $78K–$176K/yr