9 hours ago
Responsibilities
- Design, implement, and maintain infrastructure for reliable and rapid foundation model evaluation.
- Collaborate with machine learning researchers on model hillclimbing and novel model-performance measurement methodologies.
- Support model shipping decisions by assessing the capabilities and performance of models powering Apple Intelligence features.
- Work across pre-training, post-training, online, offline, automated, and human evaluation workflows.
Requirements
- At least 5 years of hands-on machine learning engineering experience, including at least 1 year working directly with large language models or generative AI.
- Bachelor's, master's, or PhD in computer science, machine learning, or a related technical field, or equivalent practical experience.
- Strong software engineering fundamentals in debugging, testing, code reviews, production reliability, and scalability.
- Preferred: experience evaluating large language models at scale or designing large language model benchmarks.
- Preferred: strong communication, creative and critical thinking, curiosity, self-motivation, and ability to prioritize problems amid ambiguity.
Categories
About Apple
Apple designs and sells consumer electronics, software, and services for consumers and professionals worldwide, including iPhone, Mac, iPad, Apple Watch, and AirPods, plus platforms like iOS/macOS and services such as the App Store, iCloud, Music, and TV+. Its business combines device sales with services and subscriptions and in-house silicon design. Founded in 1976, Apple is headquartered in Cupertino, California, and trades on NASDAQ as AAPL.
