Apple

Sr Engineering Program Manager, Evaluation - Special Projects

Apple Cupertino, California, United States

Computers and Electronics Manufacturing · 10,001+ employees

6 h ago
Principal (10+ yrs) Full-time United States
Log in to apply, save this posting, or score it against your profile with AI.

About the role

You will lead the strategic direction for AI evaluation frameworks and translate ambiguous product requirements into concrete success metrics. This role involves driving high-stakes programs and influencing cross-functional teams at the executive level to ensure AI models meet quality standards.

What they look for

Technical Program Management AI/ML Programs Large Language Models Vision-Language Models Multi-modal Architectures Evaluation Methodologies Model Development Lifecycle Reinforcement Learning From Human Feedback Preference Optimization Statistical Analysis A/B Testing Experimental Design Executive Communication Strategic Planning Cross-functional Leadership

Requirements

Candidates must have at least 7 years of technical program management experience, including 3 years leading complex AI/ML programs. A Master's or PhD in a quantitative field is preferred, along with deep expertise in LLMs, VLMs, and statistical evaluation methodologies.

Full description

Apple's Special Projects team is seeking a Senior Engineering Program Manager (EPM) to lead our AI evaluation framework at the forefront of next-generation AI experiences. This is a highly visible role where you'll drive the strategic direction for how we measure and validate AI model performance across modalities.

You will organize and lead teams to architect our evaluation methodology, translating ambiguous product requirements into concrete success metrics that determine if our models meet Apple's quality bar. Your work will directly influence product roadmaps and drive critical hill-climbing decisions that shape the AI experiences millions of users will interact with daily.

This role requires mastery of both technical depth in ML evaluation and the ability to influence across teams at the executive level. You'll lead complex, high-stakes programs while mentoring teams and establishing best practices that will define Apple's approach to AI quality.

Description

As the EPM for Evaluation, you will be responsible for :

Minimum Qualifications

7+ years of technical program management experience, with at least 3 years leading complex AI/ML programs. Proven track record of delivering complex programs by defining clear requirements and driving engineering teams to successful outcomes. Experience influencing and driving decisions at Director/VP level. Strong ability to navigate ambiguity and lead teams through uncertainty while maintaining program momentum. Excellence in executive communication - ability to distill complex technical information with the right balance of detail.

Preferred Qualifications

Master's or PhD in Computer Science, Machine Learning, Statistics, or related quantitative field. Deep hands-on experience with large language models (LLMs), vision-language models (VLMs), and multi-modal architectures. Demonstrated mastery of ML concepts, evaluation methodologies, and the end-to-end model development lifecycle. Experience with reinforcement learning from human feedback (RLHF) and preference optimization. Expertise in statistical analysis, A/B testing, and experimental design at scale.