In-House Marketing Assistant Manager - San Diego Travel + Leisure Co.In-House Marketing Assistant Manager - San DiegoSan Diego, CaliforniaTravel + Leisure Co. is the world’s leading vacation ownership and travel membership company, with a dynamic and growing portfolio of resort, travel club, and lifestyle travel brands. Responsibilities include, but are not limited to: • Direct supervision of In-House Marketing staff: interview, hire and train associates; plan, assign and direct work; conduct performance reviews; motivate, reward, and provide disciplinary action when necessary (termination and conflict resolution).
ML Scientist I / II, Foundation Models for Life Sciences Lila SciencesML Scientist I / II, Foundation Models for Life SciencesSan Francisco, CaliforniaFull-time U.S. employees receive a comprehensive benefits program including medical, dental, and vision coverage; employer-paid life and disability insurance; flexible time off with generous company wide holidays; paid parental leave; an educational assistance program; commuter benefits, including bike share memberships for office based employees; and a company subsidized lunch program. You will work on generative models spanning biological sequences, molecular structures, and multimodal experimental data, contributing to problem formulation, model design, training, evaluation, and integration into Lila's closed-loop discovery engine.
Senior / Principal ML Scientist, Foundation Models for Life Sciences Lila SciencesSenior / Principal ML Scientist, Foundation Models for Life SciencesSan Francisco, CaliforniaFull-time U.S. employees receive a comprehensive benefits program including medical, dental, and vision coverage; employer-paid life and disability insurance; flexible time off with generous company wide holidays; paid parental leave; an educational assistance program; commuter benefits, including bike share memberships for office based employees; and a company subsidized lunch program. LILA combines advanced AI models with proprietary AI Science Factory instruments into an operating system for science that executes the entire scientific method autonomously, accelerating discovery at unprecedented speed, scale, and impact across medicine, materials, and energy.
Software Engineer, Models MeterSoftware Engineer, ModelsSan Francisco, CaliforniaIn addition to your customers, network engineers, you’ll partner closely with two research engineers who have deep ML backgrounds and a clear picture of what training data needs to look like. When a network engineer looks at a set of device stats and figures out it’s upstream packet loss — not a hardware failure, not a misconfiguration, specifically upstream packet loss — that reasoning lives in their head.
Senior Research Scientist, Multimodal Foundation Models And Robotics NvidiaSenior Research Scientist, Multimodal Foundation Models And RoboticsSanta Clara, CADeep understanding of robot kinematics, dynamics, and sensors; Ability to safely operate robot hardware, lab equipment, and tools; Knowledge of control methods, including PID, model predictive control, and whole-body control; Familiarity with physics simulation frameworks such as MuJoCo and Isaac Sim; Robot hardware design and hands-on building experience. Hands-on training experience and publications in at least one of the following topics: LLMs; Large vision-language models; Video generative models and diffusion algorithms; or Action-based transformers.
Senior Manager, Interactive World Model Platforms NvidiaSenior Manager, Interactive World Model PlatformsSanta Clara, CATechnical fluency in the ML primitives behind interactive world models, including diffusion or flow-matching models, autoregressive / causal video generation, self-forcing or causal-forcing style training, Gaussian splatting, NeRFs, and neural reconstruction. Ways to stand out from the crowd: Experience adopting Gaussian splats, NeRFs, neural reconstruction, neural shading, or other advanced rendering techniques for AV, robotics, simulation, synthetic data, or production rendering workflows.
Senior Software Engineer, Spatial Intelligence And Foundation Models NvidiaSenior Software Engineer, Spatial Intelligence And Foundation ModelsSanta Clara, CAYou will join a group of world-class robotics software and applied research engineers focused on geometric and semantic understanding, and reasoning for robots - building the perception systems that turn raw sensor data into actionable world understanding, shaping the future of physical AI! What you'll be doing: Design, implement, and deploy novel algorithms for spatial understanding, working on problems ranging from SLAM, structure-from-motion, optical flow, scene flow, and object reconstruction to training VLMs on a wide range of spatial reasoning skills.
Staff Software Engineer, Modelling Infrastructure GoogleStaff Software Engineer, Modelling InfrastructureSunnyvale, CAYou will work closely with partner teams, external organizations, and engineers from various tool and infrastructure teams to understand needs, align road-maps, and ensure the modeling infrastructure meets the evolving demands of both networking and AI/ML landscape. From software to hardware our teams are shaping the future of world-leading hyperscale computing, with key teams working on the development of our TPUs, Vertex AI for Google Cloud, Google Global Networking, Data Center operations, systems research, and much more.
AI/ML Engineer - Model Inference GMAI/ML Engineer - Model InferenceSunnyvale, CaliforniaThe team sits at the intersection of machine learning, data infrastructure, and developer productivity, building systems that make it easier to search for important scenarios, prepare training-ready data, and support fast iteration across perception and evaluation workflows. We believe the next generation of autonomy and robotics depends not only on stronger models, but also on better infrastructure for turning massive volumes of multimodal data into reusable signals, searchable artifacts, and high-quality evaluation loops.
Lead Software Engineer, Model Serving Platform SciforiumLead Software Engineer, Model Serving PlatformSan Francisco, CaliforniaBacked by multi-million-dollar funding and direct sponsorship from AMD with hands-on support from AMD engineers the team is scaling rapidly to build the full stack powering frontier AI models and real-time applications. Experience with ML systems engineering, distributed GPU scheduling, open source inference engine like vLLM, Sglang, or TRT-LLM.
Machine Learning Infrastructure Engineer, Model Inference AbridgeMachine Learning Infrastructure Engineer, Model InferenceSan Francisco, CaliforniaAs an ML Infrastructure Engineer, Model Inference at Abridge, you’ll play a pivotal role in building and optimizing the core inference infrastructure that powers our machine learning models. Powered by Linked Evidence and our purpose-built, auditable AI, we are the only company that maps AI-generated summaries to ground truth, helping providers quickly trust and verify the output.
Engineering Manager, Model Library BasetenEngineering Manager, Model LibrarySan Francisco, CaliforniaYou'll lead the Model Library team at Baseten — a small, high-ownership team focused on helping developers discover, evaluate, and select the right models for their specific use cases. This role is for a product-minded engineering manager who thrives in ambiguous, high-impact environments and can guide a team from early-stage product thinking through production-quality execution.
Sr. ML Production Model Automation Engineer, Siri Speech Apple IncSr. ML Production Model Automation Engineer, Siri SpeechCupertino, CADesign and operate agent-based automation pipelines for ML models where agents own decision logic at each gate and humans approve only at defined escalation points Develop multi-agent workflows using LLM-native tooling for on-device evaluation, regression triage, release readiness decisions, and automated root cause analysis. Production experience with one or more cloud ML platforms (GCP TPU, AWS GPU clusters, Kubernetes-backed training infra) including submitting jobs, debugging schedulers, working around quota systems.
AIML - Distinguished Engineer, Foundation Models Apple IncAIML - Distinguished Engineer, Foundation ModelsCupertino, CADemonstrated expertise in deep learning with a publication record in relevant conferences (e.g., NeurIPS, ICML, ICLR, CVPR, ICCV, ECCV, KDD, ACL, ICASSP, InterSpeech) or a track record in applying deep learning techniques to products Proficient programming skills in Python and one of the deep learning toolkits such as JAX, PyTorch, or Tensorflow Ability to work in a collaborative environmentWeb-scale information retrieval Human-like conversation agent Multi-modal perception for existing products and future hardware platforms On-device intelligence and learning with strong privacy protections PhD, or equivalent practical experience, in Computer Science, or related technical field. Were solving frontier problems in reward modeling to resist reward hacking, handling sparse and delayed rewards in agentic settings, and aligning models reliably across the spectrum from open-ended creative tasks to precise, action-taking workflows.
Senior Machine Learning Engineer - Foundation Model XPeng MotorsSenior Machine Learning Engineer - Foundation ModelSanta Clara, CA$174,720–$295,680 / yearYou will work closely with world-class researchers, perception and planning engineers, and infrastructure experts to design, train, and deploy large-scale multi-modal models that unify vision, language, and control. XPENG is a leading smart technology company at the forefront of innovation, integrating advanced AI and autonomous driving technologies into its vehicles, including electric vehicles (EVs), electric vertical take-off and landing (eVTOL) aircraft, and robotics.
Research Scientist, World Models (Intelligent Creation) - Global Frontier Tech Recruitment Program - 2027 Start (PhD) TikTok IncResearch Scientist, World Models (Intelligent Creation) - Global Frontier Tech Recruitment Program - 2027 Start (PhD)San Jose, CADemonstrated ability to communicate complex technical concepts and collaborate effectively within cross-functional research teams * Proven track record of first-author publications in prestigious venues including CVPR, ICLR, NeurIPS, SIGGRAPH, and ICML. About the Team The Intelligent Creation - Global GenAI team focuses on applied research in Generative AI, and delivers intelligent solutions to TikTok, enabling users to make and share creative content in a much easier way.
AI Engineer, Model Quality and Performance Cerebras SystemsAI Engineer, Model Quality and PerformanceSunnyvale, CAYoull use AI agents to spin up custom eval suites per customer use case, mine trajectories for representative test data, automate the repetitive parts of release qual, and help build performance datasets and benchmarking workflows for customer use cases. You will define what "good" looks like across the models we serve, building AI-driven systems to measure it at scale, and translating those signals into artifacts our customers and product team actually use.
Senior / Staff AI Research Scientist, Foundation Models RoboForceSenior / Staff AI Research Scientist, Foundation ModelsMilpitas, CaliforniaIn this role, you will develop algorithms that enable robots to understand their environment, interpret and execute tasks, and communicate seamlessly with humans — with a particular focus on building and training world models that allow robots to predict, plan, and generalize across complex physical tasks. Bonus Qualifications Experience with video generation or prediction models (e.g., diffusion-based video models, autoregressive video transformers) and their application to world modeling or synthetic data generation for robot learning.
Research Scientist - World Model LumaResearch Scientist - World ModelRedwood City, CaliforniaLuma already trains the strongest generative video models in the industry; the next step is turning those models into world models — interactive, controllable, physically faithful, and useful as a substrate for embodied reasoning. WHAT YOU'LL DO - Invent next-generation world model architectures — diffusion, transformer, autoregressive, or hybrid — with a particular focus on controllability and physical consistency.
Research Scientist - World Model Luma AI IncResearch Scientist - World ModelSan Francisco, CALuma already trains the strongest generative video models in the industry; the next step is turning those models into world models - interactive, controllable, physically faithful, and useful as a substrate for embodied reasoning. Invent next-generation world model architectures - diffusion, transformer, autoregressive, or hybrid - with a particular focus on controllability and physical consistency.
World Model Research Scientist- Physical AI KodiakWorld Model Research Scientist- Physical AIMountain View, CA$190,000–$250,000 / yearShould the position require, and Kodiak determines that a candidate's residence, U.S. person status, and/or citizenship status necessitate an export license, bar the candidate from the position, or otherwise fall under national security-related restrictions, Kodiak will consider the candidate for alternative positions unaffected by such restrictions, under terms and conditions set forth at Kodiak's sole discretion, or, as an alternative, opt not to proceed with the candidate's application. We are looking for a research scientist to lead the design and development of world models capable of generating multi-sensor, multi-view, temporally coherent driving scenarios conditioned on actions, 3D scene context, and text.
Staff Software Engineer, Modelling Infrastructure Google LLCStaff Software Engineer, Modelling InfrastructureSunnyvale, CAWe"re looking for engineers who bring fresh ideas from all areas, including information retrieval, distributed computing, large-scale system design, networking and data storage, security, artificial intelligence, natural language processing, UI design and mobile; the list goes on and is growing every day. You will work closely with partner teams, external organizations, and engineers from various tool and infrastructure teams to understand needs, align road-maps, and ensure the modeling infrastructure meets the evolving demands of both networking and AI/ML landscape.
Senior Director, Device and SPICE Modeling NVIDIA CorpSenior Director, Device and SPICE ModelingSanta Clara, CAThe "Device & SPICE modeling" group is responsible to co-develop with foundries in advanced device technology in the following three areas: Achieve device performance targets in Speed, Leakage, and Variation, Release accurate SPICE model based on test chip data, Tape-out device/RO test structures in test chip for SPICE model validation, process readiness & product scribe line monitors. The Advanced Technology Group (ATG) at NVIDIA is an organization of process, CAD, and design engineers that works closely with key foundry partners and internal design groups.
Senior Director, Device And Spice Modeling NvidiaSenior Director, Device And Spice ModelingSanta Clara, CAThe "Device & SPICE modeling" group is responsible to co-develop with foundries in advanced device technology in the following three areas: Achieve device performance targets in Speed, Leakage, and Variation, Release accurate SPICE model based on test chip data, Tape-out device/RO test structures in test chip for SPICE model validation, process readiness & product scribe line monitors. The Advanced Technology Group (ATG) at NVIDIA is an organization of process, CAD, and design engineers that works closely with key foundry partners and internal design groups.
Research Engineer - Model Architectures Zyphra TechnologiesResearch Engineer - Model ArchitecturesSan Francisco, CAWe strongly value new and crazy ideas and are very willing to bet big on new ideas. The Role: As a Research Engineer - Model Architectures, you will be a core contributor to Zyphra's AI Architecture Research Team.
Research Engineer - Model Architectures ZyphraResearch Engineer - Model ArchitecturesSan Francisco, CaliforniaWe strongly value new and crazy ideas and are very willing to bet big on new ideas. Generally, a joy in inventing and seriously assessing ‘crazy’ ideas, and the ability to have a unique perspective on things.
ML Engineer, Foundation Models Humble RoboticsML Engineer, Foundation ModelsSan Francisco, CaliforniaWe’re building an autonomous, zero-emissions hauler that dramatically lowers the cost of freight with groundbreaking vision-based AI, designed for today’s global logistics network. Translate state-of-the-art research (diffusion/flow-matching action heads, reasoning-augmented VLAs, world models) into production-grade systems.
Workload / Performance Model Lead SiFive IncWorkload / Performance Model LeadSanta Clara, CA$231,444–$282,876 / yearSiFive's unrivaled compute platforms are continuing to enable leading technology companies around the world to innovate, optimize and deliver the most advanced solutions of tomorrow across every market segment of chip design, including artificial intelligence, machine learning, automotive, data center, mobile, and consumer. Any offer of employment for this position is also contingent on the Company verifying that you are a authorized for access to export-controlled technology under applicable export control laws or, if you are not already authorized, our ability to successfully obtain any necessary export license(s) or other approvals.
Model Maker Ursus, Inc.Model MakerFoster City, CA$52–$62.93 / hourYou will both bring and grow a prototyping mindset in addition to various capabilities including (but not limited to) metalwork (waterjet, shear, brake, mill, lathe, welding), composites, and additive manufacturing. Your support will include interfacing with both prototyping teammates and internal clients to support the rapid iteration for vehicle development.
Senior AI Inference Engineer - Model Optimization & Deployment ZooxSenior AI Inference Engineer - Model Optimization & DeploymentFoster City, CAWe are looking for experts with hands-on experience in compressing, accelerating, and deploying complex models (LLMs, VLMs, or FMs) for power- and thermal-constrained vehicle SOCs. You will optimize the ML models, write custom CUDA kernels, and build highly concurrent inference code to ensure real-time, deterministic execution on edge devices.
Engineering Manager, Model Library Baseten Labs IncEngineering Manager, Model LibraryCATHE ROLE: Youll lead the Model Library team at Baseten a small, high-ownership team focused on helping developers discover, evaluate, and select the right models for their specific use cases. Youll stay technically grounded while driving the teams work across model discovery, evaluation frameworks, and the infrastructure that powers a best-in-class model library experience.
NewStaff Software Engineer- Foundation Model Inference Databricks IncStaff Software Engineer- Foundation Model InferenceSan Francisco, CA$190,000–$265,000 / yearThe impact you will have: Build LLM infrastructure powering large-scale inference workloads for customers through partner models (OpenAI, Anthropic, Gemini) and self-hosted models (Qwen, GPT-OSS, Llama). More than 10,000 organizations worldwide - including Comcast, Condé Nast, Grammarly, and over 50% of the Fortune 500 - rely on the Databricks Data Intelligence Platform to unify and democratize data, analytics and AI.
Platform Engineer, Model Shaping Together AIPlatform Engineer, Model ShapingSan Francisco, CA$200,000–$290,000 / yearWe believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancements such as FlashAttention, RedPajama, SWARM Parallelism, and SpecExec.
Member of Technical Staff, Model Evaluation MirendilMember of Technical Staff, Model EvaluationSan Francisco, CaliforniaWe believe accelerating scientific discovery is one of the most powerful ways to improve the future of humanity, and that AI will play a central role in making that possible. Mirendil is a tech-first company focused on solving core bottlenecks that unlock step-change acceleration across science and technology.
Engineering Manager, Foundation Model Inference FMAPI Databricks IncEngineering Manager, Foundation Model Inference FMAPIMountain View, CA$190,000–$261,250 / yearMore than 10,000 organizations worldwide - including Comcast, Condé Nast, Grammarly, and over 50% of the Fortune 500 - rely on the Databricks Data Intelligence Platform to unify and democratize data, analytics and AI. Strong technical judgment in distributed systems, platform infrastructure, AI/ML infrastructure, or large-scale backend services, with the ability to maintain quality even in areas you have not worked on personally.
Research Scientist, Foundation Model PikaResearch Scientist, Foundation ModelPalo Alto, CaliforniaDesign and prototype novel algorithms and architectures for high-fidelity, real-time multimodal synthesis and interaction across modalities. Advance state-of-the-art techniques in diffusion, autoregressive, and other generative models for large-scale pre-training and fine-tuning.
NewAssociate, Quantitative Developer, Model Portfolio Solutions (MPS), Multi-Asset Strategies & Solutions (MASS) BlackRock IncAssociate, Quantitative Developer, Model Portfolio Solutions (MPS), Multi-Asset Strategies & Solutions (MASS)San Francisco, CA$116,000–$155,000 / yearAs a Quantitative Developer on the Model Portfolio Solutions team, you will sit at the heart of BlackRock's innovation engine for quantitatively driven investing-designing, building, and scaling the signal implementations and analytical tools that researchers and portfolio managers rely on every day. You will translate cutting-edge quantitative research into production-grade signals, pioneer the integration of AI and agentic tooling into investment workflows, strengthen investment controls, and deliver scalable solutions that shape how a global, fast-growing systematic business invests.
Coach (Pyramid Model) KidangoCoach (Pyramid Model)Fremont, CaliforniaThis position is responsible for supporting program-wide fidelity to the Pyramid Model framework, enhancing classroom practices, and building staff capacity to promote children’s social-emotional development, promote self-regulation, and implement evidence-based interventions. The Coach (Pyramid Model) provides specialized coaching, training, and technical assistance to teachers, administrators, and support staff in the implementation of the Pyramid Model for Promoting Social Emotional Competence in Infants and Young Children.
AIML Researcher/Engineer - Foundation Model Post-Training Apple IncAIML Researcher/Engineer - Foundation Model Post-TrainingCupertino, CAIn this role, you will play a critical role shaping the future of our LLM efforts, specifically in transforming our models into highly capable, intelligent assistants that power billions of Apple products. Demonstrated expertise in deep learning with a focus on LLMs, post-training, or reinforcement learning, backed by a strong record of academic or real-world accomplishments in these or closely related domains.
Senior Machine Learning Engineer - Model Inference AppleSenior Machine Learning Engineer - Model InferenceCupertino, CAYour responsibilities span the full server stack, including onboarding new use cases, optimizing inference across heterogeneous accelerated compute hardware, deploying services on Kubernetes, building and integrating inference engines and control-plane components, and ensuring seamless integration with Maps infrastructure. **Description** As a Software Engineer on the Apple Maps team, you will lead the design and implementation of large-scale, high-performance inference services that support a wide range of models used across Maps, including deep learning and large language models.
Software Engineer - Model Performance - Special Projects Apple IncSoftware Engineer - Model Performance - Special ProjectsCupertino, CAWhat the day-to-day looks like:Investigating ML model performance and power on Apple Silicon Working closely with performance, power, model, and silicon teams to optimize the system Partner with ML and AI researchers to translate cutting-edge models onto Apple Silicon Prototyping ideas to support feature definition and iteration Clearly communicate technical concepts to cross functional partners Collaboration on architecture discussions and code reviewYouve shipped software that people relied on, and you learned something from every hard call along the way Youre someone who unblocks yourself. Deep experience in one or more iOS/macOS domains: system services, UI frameworks, concurrent application architecture, or performance optimization Practical knowledge of how LLMs work at the lowest level of model design and execution Close to the frontier.
NewPartner Sales Director - AI Alliances - Model Providers Dynatrace IncPartner Sales Director - AI Alliances - Model ProvidersSan Francisco, CA$192,000–$240,000 / yearThis role operates within the AI Native Ecosystem Alliance charter and collaborating across AI native field sales, Product and Marketing to ensure AI partnerships generate partner-influenced pipeline and revenue, enterprise customer wins and new logos. This executive serves as Dynatrace''s primary alliance owner for a targeted list of prioritized partners including Anthropic OpenAI, Mistral, Groq, Cohere and other leading open source and commercial providers.
Senior Staff Machine Learning Engineer – Autonomous Driving Foundation Models XPeng MotorsSenior Staff Machine Learning Engineer – Autonomous Driving Foundation ModelsSanta Clara, CA$244,140–$413,160 / yearXPENG is a leading smart technology company at the forefront of innovation, integrating advanced AI and autonomous driving technologies into its vehicles, including electric vehicles (EVs), electric vertical take-off and landing (eVTOL) aircraft, and robotics. Key Responsibilities: Architectural Leadership: Lead the design of end-to-end VLA architectures, bridging multi-modal perception with high-level linguistic reasoning and precise action generation.
Staff Machine Learning Engineer - Foundation Model XPeng MotorsStaff Machine Learning Engineer - Foundation ModelSanta Clara, CA$215,280–$364,320 / yearDesign, train, and deploy large deep learning models that can leverage the vast amount of labeled and unlabeled data from a fleet of million vehicles. Within the range, individual pay is determined by work location and additional factors, including job-related skills, experience, and relevant education or training.
Model Designer OpenAIModel DesignerSan Francisco, CaliforniaFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity.
NewResearch Program Manager – Adversarial Model Research OpenAIResearch Program Manager – Adversarial Model ResearchSan Francisco, CaliforniaFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. The Human Data team at OpenAI is responsible for identifying and mitigating risks in advanced AI systems by designing evaluations, surfacing vulnerabilities, and collaborating closely with researchers to strengthen model reliability and public trust.
Solutions Architect - AI Model Specialist FriendliAISolutions Architect - AI Model SpecialistSan Francisco, CaliforniaOur infrastructure powers high-throughput, low-latency workloads for global organizations and integrates directly with Hugging Face, providing instant access to over 500,000 open-source models. You will work closely with our customers to integrate FriendliAI’s inference and agent frameworks into real-world products, enabling them to build and scale AI applications effectively.
Machine Learning Engineer, Model Evaluations (Speech LLM) - San Francisco PlaudMachine Learning Engineer, Model Evaluations (Speech LLM) - San FranciscoSan Francisco, CaliforniaWith a mission to amplify human intelligence, Plaud captures, structures, and compounds the intelligence generated in conversations — so humans can think better, decide faster, and execute with clarity. Can deeply partner with ML researchers to define exactly what "good" looks like for a Speech LLM, translating capabilities (like ASR robustness in noisy environments or TTS emotional steerability) into measurable benchmarks.
Software Development Manager, LLM Inference Model Enablement, Neuron SDK Amazon.com IncSoftware Development Manager, LLM Inference Model Enablement, Neuron SDKCupertino, CAAs an SDM for the LLM Inference Model Enablement team, you will lead a team of expert AI/ML engineers to onboard and optimize state-of-the-art open-source and customer LLMs, both dense and MoE, for inference on Trainium accelerators. The ideal candidate will have a strong background in LLM model architectures, model performance optimizations, and inference techniques, such as delivering high-performance models using distributed inference libraries.
Edge ML Software Engineer (Model Optimization-PICO) - San Jose Beijing ByteDance Technology Co LtdEdge ML Software Engineer (Model Optimization-PICO) - San JoseSan Jose, CAMaster's degree in Computer Science, Electrical Engineering, Computer Engineering, or a related field, or equivalent practical experience. Apply hardware-aware optimization strategies, such as quantization, compression and operator fusion, to meet latency, memory and power targets.