Apple logo

Machine Learning Engineer, Foundation Model Services

Apple

Santa Clara, CAJobSeen 3 days agoSeen in employer's feed 3 days ago

Most applications go out cold — see where you stand first. No sign-up to start.

Watch jobs like this.

At a glance

Compensation
No compensation found
Location
Santa Clara, CA
Work Authorization
Not specified

Olive lists jobs from US employers, including remote roles you can work from the United States.

Requirements

Credentials this posting asks for.

Bachelor's degree

Job overview

Apple’s Foundation Model Services team builds frameworks, services, and tools for large foundation models in production, powering intelligent experiences across Apple products such as Search, Music, TV, the App Store, Messages, Photos, Spotlight, Safari, and Siri, serving millions of queries with low latency and leveraging Apple hardware.

Skills & qualifications

RequiredNice to have

Skills

GoPythonAWSAzureGCPKubernetesDockerLLMsMachine LearningNLPInformation RetrievalStatisticsNvidia TensorRT-LLMvLLLMDeepSpeedNvidia Triton ServerStrong CommunicationCollaboration

Qualifications

Bachelor’s Degree in Computer Science or Related Technical Field2+ Years Industry Experience Building and Shipping Production Software and/or Machine Learning Systems5+ Years Industry Experience in ML Technologies

Full job description

Role Number: 200676373-3760

Summary

Do you think differently? Are you eager to break the status quo, bold and ambitious, unafraid to take risks, and passionate about building best-in-class technology? If so, there's no better place to do it than Apple. The Foundation Model Services team builds the frameworks, services, and tools that run Apple's largest foundation models in production. Our infrastructure powers intelligent experiences across products people use every day — Search, Music, TV, the App Store, Messages, Photos, Spotlight, Safari, Siri, and more — serving millions of queries at incredibly low latency while drawing every ounce of performance from our hardware. Join us and you'll help bring intelligence to billions of users around the world, working on optimizing and serving large language, vision, and speech models at Apple's scale.

Description

Work closely with product teams to build production-grade solutions that launch models serving customers in real time. Partner with foundation model researchers to prototype and develop inference for cutting-edge model architectures, and build tools that help us understand and remove performance bottlenecks across different hardware and use cases. Write high-quality code, learn quickly in a fast-moving field, and grow your impact as you take on larger pieces of the system.

Minimum Qualifications

  • 2+ years of industry experience building and shipping production software and/or machine learning systems.

  • Proficiency in a modern programming language such as Go or Python.

  • Experience deploying and operating services on a cloud platform (AWS, Azure, GCP, or equivalent) using containers and Kubernetes/Docker.

  • 5 year+ industry experience in ML technologies (LLMs, Machine Learning, NLP, Information Retrieval, Statistics).

  • Experience building or operating high-throughput, low-latency services.

  • Strong communication and collaboration skills, with the ability to partner across research and product teams.

  • Bachelor’s degree or higher in Computer Science or related technical field.

Preferred Qualifications

  • Familiarity with Nvidia TensorRT-LLM, vLLLM, DeepSpeed, Nvidia Triton Server etc.

Apple is an equal opportunity employer that is committed to inclusion and diversity. We seek to promote equal opportunity for all applicants without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, Veteran status, or other legally protected characteristics.

Similar jobs, posted recently

Open roles like this one, listed in the last 30 days.

You've read the whole posting — now see how you match it.