DevOps Engineer - AI Model Evaluator - AI Trainer MercorDevOps Engineer - AI Model Evaluator - AI TrainerSan Francisco, CaliforniaRemoteRegular use of AI coding agents like Cursor , Claude Code , Codex , Windsurf , Gemini CLI , or similar tools. Use frontier AI coding agents to complete and evaluate complex infrastructure engineering tasks.
Senior Director, Device and SPICE Modeling NVIDIA CorpSenior Director, Device and SPICE ModelingSanta Clara, CAThe "Device & SPICE modeling" group is responsible to co-develop with foundries in advanced device technology in the following three areas: Achieve device performance targets in Speed, Leakage, and Variation, Release accurate SPICE model based on test chip data, Tape-out device/RO test structures in test chip for SPICE model validation, process readiness & product scribe line monitors. The Advanced Technology Group (ATG) at NVIDIA is an organization of process, CAD, and design engineers that works closely with key foundry partners and internal design groups.
Senior Director, Device And Spice Modeling NvidiaSenior Director, Device And Spice ModelingSanta Clara, CAThe "Device & SPICE modeling" group is responsible to co-develop with foundries in advanced device technology in the following three areas: Achieve device performance targets in Speed, Leakage, and Variation, Release accurate SPICE model based on test chip data, Tape-out device/RO test structures in test chip for SPICE model validation, process readiness & product scribe line monitors. The Advanced Technology Group (ATG) at NVIDIA is an organization of process, CAD, and design engineers that works closely with key foundry partners and internal design groups.
AI Inference Engineer Intern - Model Pruning quadric, IncAI Inference Engineer Intern - Model PruningBurlingame, CA$45–$60 / hourQuadric's co-optimized software and hardware is targeted to run neural network (NN) inference workloads in a wide variety of edge and endpoint devices, ranging from battery operated smart-sensor systems to high-performance automotive or autonomous vehicle systems. Unlike other NPUs or neural network accelerators in the industry today that can only accelerate a portion of a machine learning graph, the Quadric GPNPU executes both NN graph code and conventional C++ DSP and control code.
Workload / Performance Model Lead SiFive IncWorkload / Performance Model LeadBerkeley, CA$231,444–$282,876 / yearSiFive's unrivaled compute platforms are continuing to enable leading technology companies around the world to innovate, optimize and deliver the most advanced solutions of tomorrow across every market segment of chip design, including artificial intelligence, machine learning, automotive, data center, mobile, and consumer. Any offer of employment for this position is also contingent on the Company verifying that you are a authorized for access to export-controlled technology under applicable export control laws or, if you are not already authorized, our ability to successfully obtain any necessary export license(s) or other approvals.
GNC Engineer, Advanced Modeling Xona Space SystemsGNC Engineer, Advanced ModelingBurlingame, CaliforniaWith Pulsar – the world’s most advanced PNT satellite infrastructure in Low Earth Orbit – Xona will offer a future-proof, backwards-compatible global positioning system optimized for absolute precision, superior power, and robust protection. You will build nonlinear six-degree-of-freedom simulation environments, develop flexible-body and vehicle subsystem models, and validate those models using analytical, test, and flight data.
GNC Engineer, Advanced Modeling Xona Space Systems, Inc.GNC Engineer, Advanced ModelingBurlingame, CAWith Pulsar - the world's most advanced PNT satellite infrastructure in Low Earth Orbit - Xona will offer a future-proof, backwards-compatible global positioning system optimized for absolute precision, superior power, and robust protection. You will build nonlinear six-degree-of-freedom simulation environments, develop flexible-body and vehicle subsystem models, and validate those models using analytical, test, and flight data.
Power Model Engineer SiFivePower Model EngineerSanta Clara, CA$178,848–$218,592 / yearSiFive's unrivaled compute platforms are continuing to enable leading technology companies around the world to innovate, optimize and deliver the most advanced solutions of tomorrow across every market segment of chip design, including artificial intelligence, machine learning, automotive, data center, mobile, and consumer. Any offer of employment for this position is also contingent on the Company verifying that you are a authorized for access to export-controlled technology under applicable export control laws or, if you are not already authorized, our ability to successfully obtain any necessary export license(s) or other approvals.
Staff Software Engineer, Modelling Infrastructure GoogleStaff Software Engineer, Modelling InfrastructureSunnyvale, CAFull timeYou will work closely with partner teams, external organizations, and engineers from various tool and infrastructure teams to understand needs, align road-maps, and ensure the modeling infrastructure meets the evolving demands of both networking and AI/ML landscape. From software to hardware our teams are shaping the future of world-leading hyperscale computing, with key teams working on the development of our TPUs, Vertex AI for Google Cloud, Google Global Networking, Data Center operations, systems research, and much more.
Sr. ML Production Model Automation Engineer, Siri Speech Apple IncSr. ML Production Model Automation Engineer, Siri SpeechCupertino, CADesign and operate agent-based automation pipelines for ML models where agents own decision logic at each gate and humans approve only at defined escalation points Develop multi-agent workflows using LLM-native tooling for on-device evaluation, regression triage, release readiness decisions, and automated root cause analysis. Production experience with one or more cloud ML platforms (GCP TPU, AWS GPU clusters, Kubernetes-backed training infra) including submitting jobs, debugging schedulers, working around quota systems.
AIML - Distinguished Engineer, Foundation Models Apple IncAIML - Distinguished Engineer, Foundation ModelsCupertino, CADemonstrated expertise in deep learning with a publication record in relevant conferences (e.g., NeurIPS, ICML, ICLR, CVPR, ICCV, ECCV, KDD, ACL, ICASSP, InterSpeech) or a track record in applying deep learning techniques to products Proficient programming skills in Python and one of the deep learning toolkits such as JAX, PyTorch, or Tensorflow Ability to work in a collaborative environmentWeb-scale information retrieval Human-like conversation agent Multi-modal perception for existing products and future hardware platforms On-device intelligence and learning with strong privacy protections PhD, or equivalent practical experience, in Computer Science, or related technical field. Were solving frontier problems in reward modeling to resist reward hacking, handling sparse and delayed rewards in agentic settings, and aligning models reliably across the spectrum from open-ended creative tasks to precise, action-taking workflows.
Senior Machine Learning Engineer - Foundation Model XPeng MotorsSenior Machine Learning Engineer - Foundation ModelSanta Clara, CA$174,720–$295,680 / yearYou will work closely with world-class researchers, perception and planning engineers, and infrastructure experts to design, train, and deploy large-scale multi-modal models that unify vision, language, and control. XPENG is a leading smart technology company at the forefront of innovation, integrating advanced AI and autonomous driving technologies into its vehicles, including electric vehicles (EVs), electric vertical take-off and landing (eVTOL) aircraft, and robotics.
Model Designer OpenAIModel DesignerSan Francisco, CAFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity.
Research Program Manager - Adversarial Model Research OpenAIResearch Program Manager - Adversarial Model ResearchSan Francisco, CAFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. The Human Data team at OpenAI is responsible for identifying and mitigating risks in advanced AI systems by designing evaluations, surfacing vulnerabilities, and collaborating closely with researchers to strengthen model reliability and public trust.
Surfaces Model UX Designer, GeminiApp, DeepMind Google LLCSurfaces Model UX Designer, GeminiApp, DeepMindSan Francisco, CAAs a key member of the Gemini App Model UX team, you will collaborate with product managers, engineers, and research scientists to solve unique multimodal and voice issues for a highly visible representative experience of Gemini. Partner with AI product designers to craft leading-edge multimodal solutions, ensuring model responses, device signals, and screen relevancy work in concert to deliver cohesive, glanceable experiences.
ML Research Engineer, Foundation Models (Senior / Staff / Principal) Genesis TherapeuticsML Research Engineer, Foundation Models (Senior / Staff / Principal)San Mateo, CAOur generative and predictive AI platform, GEMS (Genesis Exploration of Molecular Space), integrates AI and physics into industry-leading models to generate and optimize drug molecules, including the breakthrough generative diffusion model Pearl for structure prediction. Bridge machine learning research and computational chemistry workflows, working closely with computational chemists, structural biologists, and medicinal chemists to ensure models translate effectively into real drug discovery programs.
NewStudent Researcher (Seed - LLM - Model) - 2026 Start (PhD) Beijing ByteDance Technology Co LtdStudent Researcher (Seed - LLM - Model) - 2026 Start (PhD)San Jose, CAPre-trained basic technologies, including efficient training and encapsulated deployment services, NLP, CV, video, MultiModal Machine Learning and other related pre-trained models and their downstream applications are preferred. Candidates who have published papers in accredited academic conferences, and have achieved excellent results in competitions in the fields of MultiModal Machine Learning, Computer Vision, or Machine Learning are preferred.
NewMachine Learning Engineer, Model Evaluations (Speech LLM) - San Francisco PlaudMachine Learning Engineer, Model Evaluations (Speech LLM) - San FranciscoSan Francisco, CaliforniaWith a mission to amplify human intelligence, Plaud captures, structures, and compounds the intelligence generated in conversations — so humans can think better, decide faster, and execute with clarity. Can deeply partner with ML researchers to define exactly what "good" looks like for a Speech LLM, translating capabilities (like ASR robustness in noisy environments or TTS emotional steerability) into measurable benchmarks.
NewSr. Software Development Engineer AI/ML, Inference Model Enablement, AWS Neuron Amazon.com IncSr. Software Development Engineer AI/ML, Inference Model Enablement, AWS NeuronCupertino, CASoftware Development Engineer on the Inference Model Enablement team, you will onboard and optimize state-of-the-art open-source and customer LLMs, both dense and MoE, for inference on Trainium accelerators. You"ll collaborate with cross-functional teams of applied scientists, system engineers, and product managers to architect and deliver state-of-the-art inference capabilities.
NewSoftware Development Engineer, AI/ML, AWS Neuron, Model Inference Amazon.com IncSoftware Development Engineer, AI/ML, AWS Neuron, Model InferenceCupertino, CAThe Annapurna Labs team at Amazon Web Services (AWS) builds AWS Neuron, the software development kit used to accelerate deep learning and GenAI workloads on Amazon's custom machine learning accelerators, Inferentia and Trainium. Participate in all stages of the ML system development lifecycle including distributed computing based architecture design, implementation, performance profiling, hardware-specific optimizations, testing and production deployment.
NewLarge Language Model Inference System Engineer Graduate (Applied Machine Learning) - 2027 Start Beijing ByteDance Technology Co LtdLarge Language Model Inference System Engineer Graduate (Applied Machine Learning) - 2027 StartSan Jose, CAReduce large model inference costs through system-level approaches such as disaggregated multi-role inference, distributed KV Cache systems, heterogeneous inference, elastic computing, and multi-tenant co-located inference. Volcano Ark hosts Doubao and leading industry large models, and supports enterprise AI adoption through stable, secure, and trusted solutions as well as professional algorithm and technical services.
Model UX Writer, Devices and Services Google LLCModel UX Writer, Devices and ServicesMountain View, CATeams across this area research, design, and develop new technologies to make our user"s interaction with computing faster and more seamless, building innovative experiences for our users around the world. The Platforms and Devices team encompasses Google"s various computing software platforms across environments (desktop, mobile, applications), as well as our first party devices and services that combine the best of Google AI, software, and hardware.
Software Development Manager, LLM Inference Model Enablement, Neuron SDK Amazon.com IncSoftware Development Manager, LLM Inference Model Enablement, Neuron SDKCupertino, CAAs an SDM for the LLM Inference Model Enablement team, you will lead a team of expert AI/ML engineers to onboard and optimize state-of-the-art open-source and customer LLMs, both dense and MoE, for inference on Trainium accelerators. The ideal candidate will have a strong background in LLM model architectures, model performance optimizations, and inference techniques, such as delivering high-performance models using distributed inference libraries.
Edge ML Software Engineer (Model Optimization-PICO) - San Jose Beijing ByteDance Technology Co LtdEdge ML Software Engineer (Model Optimization-PICO) - San JoseSan Jose, CAMaster's degree in Computer Science, Electrical Engineering, Computer Engineering, or a related field, or equivalent practical experience. Apply hardware-aware optimization strategies, such as quantization, compression and operator fusion, to meet latency, memory and power targets.
Software Engineer, Scientific Models Benchling IncSoftware Engineer, Scientific ModelsSan Francisco, CAProjects you might work on include: adding new models as soon as theyre published, improving model performance and scalability, and enabling scientists to automate their in-silico workflows by chaining models together into pipelines. AlphaFold or Boltz2) predict structures, predict scientific properties, and generate new drug designs, acting as a design partner and a major time saver to scientists who are creating life-saving therapeutics.
Research Program Manager – Adversarial Model Research OpenAIResearch Program Manager – Adversarial Model ResearchSan Francisco, CaliforniaFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. The Human Data team at OpenAI is responsible for identifying and mitigating risks in advanced AI systems by designing evaluations, surfacing vulnerabilities, and collaborating closely with researchers to strengthen model reliability and public trust.
Product Manager, Model Economics and Capacity Strategy Google LLCProduct Manager, Model Economics and Capacity StrategySunnyvale, CAYou will translate how efficiently our models run into how we price, package, and govern access to scarce compute, designing consumption offerings, commitment models, and PayGo/quota policy that balance customer value against physical capacity constraints. Our team works closely with creative engineers, designers, marketers, etc. to help design and develop technologies that improve access to the world"s information.
Technical Program Manager, GenAI Model Release, Google Search Google LLCTechnical Program Manager, GenAI Model Release, Google SearchMountain View, CASearch Technical Program Managers (TPgMs) serve as strategic program leads and thought partners who establish clear, ambitious, and measurable goals with strict accountability and timelines. In this role, you will partner directly with Search executives, engineering model captains, area leads, and the DeepMind OR team to guide model release evaluations and implementation.
Research Scientist in Large Language Model (LLM) - Seed - Graduates - 2027 Start (PhD) Beijing ByteDance Technology Co LtdResearch Scientist in Large Language Model (LLM) - Seed - Graduates - 2027 Start (PhD)San Jose, CAMinimum Qualifications: Currently pursuing a PhD in computer science, mathematics, engineering, or a related field, with an expected graduation date in 2027 and the ability to commit to an onboarding date by the end of 2027. Our areas of focus include model pretraining, posttraining, inference, memory capabilities, learning, interpretability and other related directions.
Research Scientist in Large Language Model (LLM) - Seed - Graduates - 2027 Start (BS/MS) Beijing ByteDance Technology Co LtdResearch Scientist in Large Language Model (LLM) - Seed - Graduates - 2027 Start (BS/MS)San Jose, CAMinimum Qualifications: Currently pursuing a Bachelor's or Master's degree in computer science, mathematics, engineering, or a related field, with an expected graduation date in 2027 and the ability to commit to an onboarding date by the end of 2027. Our areas of focus include model pretraining, posttraining, inference, memory capabilities, learning, interpretability and other related directions.
Staff AI Research Engineer, Large User Models Google LLCStaff AI Research Engineer, Large User ModelsMountain View, CAWe"re looking for engineers who bring fresh ideas from all areas, including information retrieval, distributed computing, large-scale system design, networking and data storage, security, artificial intelligence, natural language processing, UI design and mobile; the list goes on and is growing every day. From software to hardware our teams are shaping the future of world-leading hyperscale computing, with key teams working on the development of our TPUs, Vertex AI for Google Cloud, Google Global Networking, Data Center operations, systems research, and much more.
Solutions Architect - AI Model Specialist FriendliAISolutions Architect - AI Model SpecialistSan Francisco, CaliforniaOur infrastructure powers high-throughput, low-latency workloads for global organizations and integrates directly with Hugging Face, providing instant access to over 600,000 open-source models. You will work closely with our customers to integrate FriendliAI’s inference and agent frameworks into real-world products, enabling them to build and scale AI applications effectively.
Associate Manager, Production Line Maintenance, Model 3, Body in White Tesla IncAssociate Manager, Production Line Maintenance, Model 3, Body in WhiteFremont, CA$128,000–$192,000 / yearThey must be capable of collaborating with cross-functional teams and department leaders to address concerns and issues that affect part quality, production efficiency, work cell ergonomics, equipment maintenance, etc. 5+ years of extensive experience in supervision and leadership, with troubleshooting and correcting industrial equipment maintenance - breakdowns and failures through Root Cause Analysis and effective Corrective Action implementation.
Professional Engineering Trainee, PEAK Program, Model 3 (Summer 2026) Tesla IncProfessional Engineering Trainee, PEAK Program, Model 3 (Summer 2026)Fremont, CA$71,120–$106,680 / yearThe PEAK (Professional Engineering Accelerated Knowledge) Program provides early career engineers with the opportunity to develop their engineering skills in a challenging, fast-paced, and impactful manufacturing engineering role onsite at Tesla's Fremont, California facility. Execute Engineering Projects: Using advanced project management and statistical techniques taught in this program, you will bring significant impact to your engineering team and present your results to factory leadership.
Supervisor, Production Line Maintenance, Model 3, General Assembly (Night Shift) Tesla IncSupervisor, Production Line Maintenance, Model 3, General Assembly (Night Shift)Fremont, CA$108,000–$162,000 / yearThey must be capable of collaborating with cross-functional teams and department leaders to address concerns and issues that affect part quality, production efficiency, work cell ergonomics, equipment maintenance, etc. Establish systems to ensure that preventative maintenance is 100% on-time and effective for all manufacturing tools and equipment, including robots and other automated and semi-automated manufacturing systems.
Battery Modeling Engineer Zipline International IncBattery Modeling EngineerSouth San Francisco, CA$130,000–$165,000 / yearIn this role, you will own end-to-end battery modeling capabilities spanning cell-level electrochemical models (SPM/SPMe/P2D) and equivalent circuit models, pack-level battery models across multiple fidelities, and battery state estimation algorithms. Develop, validate, and own cell-level battery models across multiple fidelities, including electrochemical models (SPM/SPMe/P2D) and equivalent circuit models, to support cell selection and optimize performance, fast charging capability, and battery lifetime.
Research Scientist / Engineer - Foundation Model: Core Research Luma AI IncResearch Scientist / Engineer - Foundation Model: Core ResearchPalo Alto, CAUnified Modeling & Efficiency Drive the core research that powers all of Lumas products - co-designing multimodal representations, advancing core algorithms for long-context training, and establishing rigorous scaling laws to predict performance across compute budgets. This role offers the chance to bridge frontier research with magical, shipped products like Dream Machine and Ray3, solving novel problems where no playbook exists.
Staff Machine Learning Engineer - Agentic Models LLM RAG GenAI Eightfold AI IncStaff Machine Learning Engineer - Agentic Models LLM RAG GenAISanta Clara, CADevelop and integrate large language models (LLMs) and other state-of-the-art AI techniques to enhance agent autonomy and intelligence. Leverage enterprise data, market data, and user interactions to build intelligent and personalized agent experiences.
Associate Product Manager, Hive Models Castle Global IncAssociate Product Manager, Hive ModelsSan Francisco, CA$90,000–$120,000 / yearAs an Associate Product Manager on our Hive Models team, you will work cross-functionally with stakeholders to help define product requirements and support implementation efforts, collaborating closely with our Machine Learning, Core Infrastructure, and Product teams. Support our Hive Models portfolio by learning the needs of the product development teams, and contributing to the creation of our deep learning models that provide human-like interpretation of video, image, audio, and text.
Senior System Software Engineer, Interactive World Models NVIDIA CorpSenior System Software Engineer, Interactive World ModelsSanta Clara, CACollaborate with research, simulation, rendering, robotics, and autonomous-vehicle teams to make technical tradeoffs, integrate capabilities into downstream workflows, and own complex modules or cross-team projects that shape technical direction and roadmaps. Demonstrated engineering judgment and a record of turning technically complex prototypes into reliable products through thoughtful tradeoffs, testing, and delivery, along with ownership of complex modules or cross-team projects and clear communication with partners.
Alibaba Cloud Intelligence-International GenAI Marketing Manager-Model Commercialization Growth (Sunnyvale) Alibaba Group Holding LtdAlibaba Cloud Intelligence-International GenAI Marketing Manager-Model Commercialization Growth (Sunnyvale)Sunnyvale, CA$121,000–$198,000 / yearVoice of Market to Product: Serve as the critical bridge between market and internal teams-systematically capturing, structuring, and prioritizing customer feedback and market signals to directly inform product roadmaps, pricing, and launch timing, enabling an agile "use-to-build, sell-to-scale" operating model. Own the full commercialization lifecycle: design and operate end-to-end data tracking from lead API sign-up token consumption, using attribution models to measure and optimize conversion efficiency at every stage-turning insights into action to accelerate token growth and business outcomes.
Senior AI Architect, Foundation Models and SoC Co-Design - Autonomous Vehicles NVIDIA CorpSenior AI Architect, Foundation Models and SoC Co-Design - Autonomous VehiclesSanta Clara, CAYou will work with world-class AI researchers, silicon architects, and AV platform teams to identify the AI workloads that will define the next decade - and ensure NVIDIA platforms are architected to lead them. We are looking for a Senior AI Architect to help define the next generation of AI model paradigms for autonomous vehicles and shape how those models co-evolve with NVIDIA's future embedded SoC architectures.
Marketing Manager, Institute of Foundation Models (IFM) Institute of Foundation ModelsMarketing Manager, Institute of Foundation Models (IFM)Sunnyvale, CaliforniaStrategic and innovative problem-solving skills will be instrumental in establishing MBZUAI as a global hub for high-performance computing in deep learning, driving impactful discoveries that inspire the next generation of AI pioneers. Based in Sunnyvale, this role focuses on translating complex technical work into clear, compelling campaigns, content, and assets for developers, researchers, and the broader AI ecosystem.
NewProduct Manager, Growth, Scientific AI Models BenchlingProduct Manager, Growth, Scientific AI ModelsSan Francisco, CaliforniaOwn the product roadmap for our scientific model platform — defining and prioritizing the models, features, and capabilities that let scientists run inference fast, reliably, and cost-effectively across a diverse set of scientific models. AlphaFold or Boltz2) predict structures, predict scientific properties, and generate new drug designs, acting as a design partner and a major time saver to scientists who are creating life-saving therapeutics.
Product Manager - Open Models NVIDIA CorpProduct Manager - Open ModelsSanta Clara, CASupporting product marketing in crafting model-specific messaging and documentation - including blogs, whitepapers, case studies, and presentations, highlighting NVIDIA's optimizations and software tools. Collaborating closely with internal engineering, infrastructure, and developer experience teams to enable best-in-class support for open models across inference, fine-tuning, and evaluation.
Senior Technical Program Manager, GenAI and Models NVIDIA CorpSenior Technical Program Manager, GenAI and ModelsSanta Clara, CAExperience running software releases across repositories, dependencies, test configurations, quality gates, collaborator approvals, open-source workflows, CI/CD systems, and tools such as GitHub, Git, Jira, Linear, Aha!, or Confluence. NVIDIA's Deep Learning Software team is looking for a Senior Technical Program Manager to lead programs across model pre-training, production RL runs, evaluation, and agentic AI infrastructure.
Senior AI Architect, Foundation Models And Soc Co-Design - Autonomous Vehicles NvidiaSenior AI Architect, Foundation Models And Soc Co-Design - Autonomous VehiclesSanta Clara, CAYou will work with world-class AI researchers, silicon architects, and AV platform teams to identify the AI workloads that will define the next decade - and ensure NVIDIA platforms are architected to lead them. We are looking for a Senior AI Architect to help define the next generation of AI model paradigms for autonomous vehicles and shape how those models co-evolve with NVIDIA's future embedded SoC architectures.
Senior Research Engineer, Foundation Model Training Infrastructure NvidiaSenior Research Engineer, Foundation Model Training InfrastructureSanta Clara, CAWays to stand out from the crowd: Master's or PhD's degree in Computer Science, Robotics, Engineering, or a related field; Demonstrated Tech Lead experience, coordinating a team of engineers and driving projects from conception to deployment; Strong experience at building large-scale LLM and multimodal LLM training infrastructure; Contributions to popular open-source AI frameworks or research publications in top-tier AI conferences, such as NeurIPS, ICRA, ICLR, CoRL. What we need to see: Bachelor's degree in Computer Science, Robotics, Engineering, or a related field; 10+ years of full-time industry experience in large-scale MLOps and AI infrastructure; Proven experience designing and optimizing distributed training systems with frameworks like PyTorch, JAX, or TensorFlow.
Software Engineer- Model Performance Systems BaseTenSoftware Engineer- Model Performance SystemsSan Francisco, CABy uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. You will not just be building the automated "speedometer and diagnostic" suite for our next-generation AI infrastructure; you will be defining the roadmap, driving key technical decisions, and taking full ownership of the future of this work.
Senior Machine Learning Engineer - Model Inference AppleSenior Machine Learning Engineer - Model InferenceCupertino, CAYour responsibilities span the full server stack, including onboarding new use cases, optimizing inference across heterogeneous accelerated compute hardware, deploying services on Kubernetes, building and integrating inference engines and control-plane components, and ensuring seamless integration with Maps infrastructure. **Description** As a Software Engineer on the Apple Maps team, you will lead the design and implementation of large-scale, high-performance inference services that support a wide range of models used across Maps, including deep learning and large language models.