Apple logo

Senior Machine Learning Engineer - Model Inference

Apple

  • Cupertino, CA
  • 27 days ago

    Highlights

    Your responsibilities span the full server stack, including onboarding new use cases, optimizing inference across heterogeneous accelerated compute hardware, deploying services on Kubernetes, building and integrating inference engines and control-plane components, and ensuring seamless integration with Maps infrastructure. **Description** As a Software Engineer on the Apple Maps team, you will lead the design and implementation of large-scale, high-performance inference services that support a wide range of models used across Maps, including deep learning and large language models.

    Numbers & Facts

    LocationCupertino, CA
    IndustryComputer/IT Services
    Company Size10,000 employees or more
    Year Founded1976
    Websitehttps://www.apple.com/jobs

    Description

    **Weekly Hours:** 40 **Role Number:** 200670871-0836 **Summary** Join Apple Maps to help build the best map in the world. In this role on ML Platform, you will help bring advanced deep learning and large language models into high-volume, low-latency, highly available production serving, improving search quality and powering experiences across Maps. You will partner closely with research and product teams, take end-to-end ownership, and deliver measurable results at global scale. **Description** As a Software Engineer on the Apple Maps team, you will lead the design and implementation of large-scale, high-performance inference services that support a wide range of models used across Maps, including deep learning and large language models. You will collaborate closely with research and product partners to bring models into production, with a strong focus on efficiency, reliability, and scalability. Your responsibilities span the full server stack, including onboarding new use cases, optimizing inference across heterogeneous accelerated compute hardware, deploying services on Kubernetes, building and integrating inference engines and control-plane components, and ensuring seamless integration with Maps infrastructure. **Minimum Qualifications** + Bachelor's degree in Computer Science, Engineering, or related field (or equivalent experience). + 5+ years in software engineering focused directly on ML inference, GPU acceleration, and large-scale systems. + Expertise in deploying and optimizing LLMs for high-performance, production-scale inference. + Proficiency in Python, Java or C++. + Experience with deep learning frameworks like PyTorch, TensorFlow, and Hugging Face Transformers. + Experience with model serving tools (e.g., NVIDIA Triton, TensorFlow Serving, VLLM, etc) + Experience with optimization techniques like Attention Fusion, Quantization, and Speculative Decoding. + Skilled in GPU optimization (e.g., CUDA, TensorRT-LLM, cuDNN) to accelerate inference tasks. + Skilled in cloud technologies like Kubernetes, Ingress, HAProxy for scalable deployment. **Preferred Qualifications** + Master’s or PhD in Computer Science, Machine Learning, or a related field. + Understanding of ML Ops practices, continuous integration, and deployment pipelines for machine learning models. + Familiarity with model distillation, low-rank approximations, and other model compression techniques for reducing memory footprint and improving inference speed. + Strong understanding of distributed systems, multi-GPU/multi-node parallelism, and system-level optimization for large-scale inference.

    About Company

    We bring amazing people together to make amazing things happen.

    We’re a diverse collection of thinkers and doers, continually reimagining what’s possible to help us all do what we love in new ways. The people who work here have reinvented entire industries with the Mac, iPhone, iPad, and Apple Watch, as well as with services, including iTunes, the App Store, Apple Music, and Apple Pay. And the same passion for innovation that goes into our products also applies to our practices — strengthening our commitment to leave the world better than we found it.

    About Apple

    There’s a place here for every kind of brilliant. Everyone here is an innovator, or an innovator-to-be, no matter what your team or your role. So bring your passion, courage, and original thinking and get ready to share it, because every new product, service, or feature we invent is the result of people working together to make each others’ ideas stronger. Innovation at this level depends on people who represent the variety of the human experience and inspire us with their own fresh perspectives. Together, we’ll do amazing work that can make a difference in people’s lives. Including your own. Learn more about working at Apple.

    Similar Jobs