Apple logo

Staff/Sr. Machine Learning Engineer, Foundation Models - AI, Search & Knowledge Platforms

Apple
15 days ago
Full-time
On-site
Seattle, California, United States
Machine Learning
Do you feel you think differently, you are eager to break status quo, are bold and ambitious, aren’t afraid to take risks and are passionate to build the best of class technology. If yes, what better place to be at and do this than Apple? At Apple, “we think different, we push the boundaries of computing and intelligence. We build products that bring smile to people’s face”.\\nFoundation Model Infrastructure team, within AI, Search \u0026 Knowledge Platforms Technologies organization is the back-bone of Apple Intelligence. It builds frameworks, services and tools that power the largest Apple foundation models on servers. Our Infrastructure powers a wide gamut of services at Apple including Apple Search, Apple Music, AppleTV, AppStore, iMessages, Photos \u0026 Camera, Spotlight, Safari, Siri and upcoming ever exciting Apple products serving millions of queries every day with incredible low latencies, drawing every ounce of compute from our hardware.\\nAs part of this group, you will get a chance to bring Intelligence to billions of users across the world. You will have an opportunity to make a difference in life of people. You will have a chance to work on optimizing billions of parameter languge and vision and speech models using state of the art technologies and make it run at scale of Apple.

Work along side Foundation Model Research team to optimize inference for cutting edge model architectures.\\nWork closely with product teams to build Production grade solutions to launch models serving millions of customers in real time.\\nBuild tools to understand bottlenecks in Inference for different hardwares and use cases.\\nMentor and guide engineers in the organization.

7+ years of experience leading and driving complex, ambiguous projects.\\nHave experience with high throughput services particularly at supercomputing scale.\\nProficient with running applications on Cloud (AWS / Azure or equivalent) using Kubernetes, Docker etc.\\nFamiliar with GPU programming concepts using CUDA.\\nFamiliar with one of the popular ML Frameworks like Pytorch, Tensorflow.\\nBS in Computer Science, Artificial Intelligence, Machine Learning, Information Retrieval, Data Science or related field

Proficient in building and maintaining systems written in modern languages (eg: Golang, python)\\nFamiliar with fundamental Deep Learning architectures such as Transformers, Encoder/Decoder models.\\nFamiliarity with Nvidia TensorRT-LLM, vLLM, DeepSpeed, Nvidia Triton Server etc.\\nExperience writing custom CUDA kernels using CUDA or OpenAI Triton.\\nMS in Computer Science, Artificial Intelligence, Machine Learning, Information Retrieval, Data Science or related field.