Machine Learning Engineer, Evaluation, Agentic Search Capabilities
Apple Cupertino, California, United States
Computers and Electronics Manufacturing · 10,001+ employees
Applying here? Try the free cover letter tool — paste this posting and your résumé, no account needed.
About the role
You will design and build scalable systems to evaluate personalized agentic search applications across Apple products. This involves developing data generation methodologies and defining both online and offline metrics to ensure high-quality search experiences.
What they look for
Requirements
Candidates must have at least 6 years of industry experience in building machine learning or evaluation systems at scale. Proficiency in Python or C/C++ and a degree in Computer Science, Engineering, or Statistics are required.
Full description
The Agentic Search team is redefining how hundreds of millions of people access information across Apple devices, with privacy built in from the ground up. As part of the Agentic Search Evaluation team, you will help advance Apple Intelligence through personalized search that powers experiences across Siri, Spotlight, Mail, Messages, and more. Your team researches and builds deep search systems for Personal Question Answering, enabling Siri to answer questions about a user's emails, messages, events, files, and more, while keeping personal data private.
Description
We are looking for a senior Machine Learning Engineer with a passion for building scalable, high quality solutions to evaluate personalized agentic search applications. In this role, you will directly impact how we assess and improve the quality of novel agentic search applications by driving solutions for data generation, data quality assessment, offline and online metrics, and methods for statistical model and search quality assessment. Your work will impact the quality of innovative Personalized Agentic Search experiences across Siri and Apple products.
In this role you will have the opportunity to build scalable systems for comprehensive evaluation of personalized model and search systems, develop data generation and curation methodologies to support evaluations, and design online and offline metrics for product quality assessment. Our team is responsible for models and evaluation of search experiences that answer user’s questions using their personal documents with privacy at the forefront.
Minimum Qualifications
6+ years industry experience in building Machine Learning or ML evaluation systems at scale. Strong software engineering skills in mainstream programming languages, such as: Python, C/C++. Ability to quickly prototype ideas and solutions, and perform critical analysis. Strong communication skills and ability to drive solutions in collaboration with partner teams. Bachelors in Computer Science, Engineering, Statistics or related field.
Preferred Qualifications
Experience in design and building production ML systems and applications in search, NLP, recommendation systems, or information retrieval. Experience in data collection, data generation and/or data quality assessment of language, image or multi-modal data. Strong skills for quality metrics development; interpretation of evaluations; and presentation to executive audience. Advanced degree (Master’s or Ph.D.) in Computer Science, Engineering, Statistics, or related field, or equivalent industry work experience.
Similar roles
-
Staff Machine Learning and Bioinformatics Scientist (Multi-Cancer Early Detection)
Natera United States · $163K–$204K/yr
-
Senior Machine Learning Engineer
Encora Perímetro Urbano Santiago de Cali, Valle del Cauca, Colombia
-
Machine Learning Engineer
Bluetab, an IBM Company Mexico City, Mexico City, Mexico
-
Lead Technical Product Manager — Machine Learning Platform
N26 Barcelona, Catalonia, Spain
-
Senior Machine Learning Engineer
Care Access United States · $140K–$190K/yr
-
Senior Machine Learning Programmer
Epic Games London, England, United Kingdom · $233K–$341K/yr