Sr. Machine Learning Engineer, Foundation Models Inference - Cloud OS & Inference
Sign up free to see how well your resume matches this role.
What you'll do
- Build a highly performant, secure, and private inference stack for foundation models.
- Power Siri AI, Apple Intelligence, and various Apple applications with large foundation models.
- Optimize language, vision, and speech models with billions of parameters.
- Ensure low-latency delivery of AI services at massive scale.
- Extract maximum compute efficiency from underlying hardware.
What they're looking for
- Experience in building high-performance inference stacks
- Experience optimizing language, vision, or speech models
- Experience with large-scale distributed systems
- Proficiency in low-latency system design
Summarised by NextRaise from the employer’s description, which follows in full below.
Full description from employer
We are the Foundation Model Inference team within Cloud OS and AI Inference organization. We are on a mission to build the most highly performant, secure and private inference stack that powers Siri AI, Apple Intelligence and Apps that are powered with the largest foundation models.
Our systems serve billions of queries daily across Siri AI, Apple Intelligence, Apple Search, Apple Music, Apple TV, App Store, iMessage, Photos, Camera, Spotlight & Safari, at remarkably low latency with every ounce of compute extracted from the hardware beneath them. We optimize language, vision, and speech models with billions of parameters using state-of-the-art techniques and ship them at Apple scale.
This is a rare opportunity to directly shape how AI reaches billions of people worldwide.
Company
Company facts come from this company's own listings. We only show what the postings themselves carry.