Apple

AIML - Senior Machine Learning Research Engineer, LLM Post-training (Multilinguality)

Apple Cupertino, California, United States

Computers and Electronics Manufacturing · 10,001+ employees

6 h ago
machine-learning Mid (2-5 yrs) Full-time United States
Create a free account to apply — email only, no card. You can also save this posting or score it against your profile with AI.

About the role

You will develop and deploy advanced machine learning models for text language identification across diverse scripts and regions. The role involves collaborating with researchers and linguists to ensure high-quality, inclusive, and efficient language technology solutions.

What they look for

Machine Learning Natural Language Processing Python PyTorch TensorFlow JAX Neural Models Multilingual Datasets Text Classification Language Identification Embedding Models Contrastive Learning Model Optimization Hugging Face Bash

Requirements

Candidates must have at least 2 years of experience in machine learning or NLP with proficiency in Python and deep learning frameworks. A bachelor's or master's degree in a technical field is required, along with experience working with multilingual or cross-lingual datasets.

Full description

The Multilingual Intelligence team is looking for a machine learning engineer to build the next generation of text language identification systems. You will develop models that accurately detect and classify languages across diverse scripts, regions, and user contexts — forming the backbone of multilingual features used by millions of people worldwide. You will work alongside a team of world-class experts to explore novel modeling approaches, data strategies, and evaluation methodologies that push the boundaries of what's possible in language detection at scale.

Passionate about Natural Language Processing, multilingual systems, and building ML that works for everyone regardless of the language they speak? Join us to make every device fluent in every language.

Description

Text language identification is a foundational capability that powers multilingual experiences across products — from translation and search to content recommendation and accessibility. As the number of supported languages grows and user expectations rise, the challenge is building models that are not only accurate but also fair, robust, and efficient across the full spectrum of the world's languages.

We are looking for a machine learning engineer passionate about building high-quality, inclusive language technology. As a member of the team, you will work across the full model lifecycle — from data curation and model design to evaluation and deployment. You will collaborate with researchers, engineers, and linguists to ensure our language identification systems perform reliably for users everywhere, regardless of how they write or what language they use.

The successful candidate should be a strong team player with excellent oral and written communication skills and a genuine passion for building ML systems that work for the world's linguistic diversity.

Minimum Qualifications

2+ year experience in machine learning, NLP, or related fields, with hands-on experience training and evaluating neural models. Proficient programming skills in Python and at least one major deep learning framework such as PyTorch, TensorFlow, or JAX. Bachelor's or Master's degree, or equivalent practical experience, in Computer Science, Machine Learning, Computational Linguistics, or a related technical field. Experience working with multilingual, multi-script, or cross-lingual datasets.

Preferred Qualifications

Experience with text classification, language identification, dialect identification, or similar NLP tasks involving multiple languages. Familiarity with embedding models, sentence representations, or contrastive learning methods. Understanding of model optimization techniques such as quantization, pruning, or knowledge distillation, and a general interest in efficient inference. Experience with low-resource languages, code-switching, or mixed-language input handling. Familiarity with Hugging Face ecosystem (transformers, tokenizers, datasets) and modern NLP pipelines. Ability to formulate a problem, design experiments, and implement end-to-end solutions in Python and Bash. Strong communication skills and a passion for working cross-functionally across Research, Engineering, and Product teams.

Similar roles