Software Engineer, Productivity - Model Performance OpenAI LLCSoftware Engineer, Productivity - Model PerformanceSan Francisco, CAFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity.
Solutions Architect, LLM Model Builder NvidiaSolutions Architect, LLM Model BuilderSanta Clara, CA5+ years of relevant experience working with LLMs, VLMs, and large-scale inference systems, with hands-on expertise in fine-tuning, benchmarking, evaluation, optimization, and production deployment as a Research Engineer, Deep Learning Engineer, or equivalent. Develop reference architectures, playbooks, benchmark recipes, TCO calculators, and sizing models across CUDA, NeMo, Nemotron, Dynamo, TensorRT-LLM, Triton, NIMs, and related tooling.
Technical Program Manager, Model Alignment and Deployment Character.AITechnical Program Manager, Model Alignment and DeploymentRedwood City, CaliforniaTogether, these groups are responsible for transforming powerful pretrained language models into intelligent, engaging, safely aligned, and highly scalable products—working across data, compute, algorithms, infrastructure, and user insights to improve model performance and ensure reliable delivery. Program ownership: Lead planning and execution of cross-functional programs spanning data collection, annotation pipelines, alignment workflows (RLHF, DPO, Constitutional AI), safety guardrails (adversarial testing, red-teaming), and model serving.
Technical Marketing Engineer, World Models - AV Physical AI NvidiaTechnical Marketing Engineer, World Models - AV Physical AISanta Clara, CAPrototype and iterate rapidly on experiments across powerful AI domains, including agentic systems, reinforcement learning, reasoning, and video generation in partnership with customer / partner teams. We are seeking a hardworking and technically skilled engineer to join our Physical AI Technical Marketing team, focusing on building world class technical materials for world models within various industries.
Engineering Manager, Model Flywheel OpenAI LLCEngineering Manager, Model FlywheelSan Francisco, CAFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. The ChatGPT Model Capabilities and Deployment team unified goal is to transform model advancements into great ChatGPT user experiences through reliable serving, rapid experimentation, safe deployment, and continuous improvement.
Senior AI/ML Research Engineer - Model development Intuitive Surgical IncSenior AI/ML Research Engineer - Model developmentSunnyvale, CACertain information you provide as part of the application will be used for purposes of determining whether Intuitive Surgical will need to (i) obtain an export license from the U.S. Government on your behalf (note: the government's licensing process can take 3 to 6+ months) or (ii) implement a Technology Control Plan ("TCP") (note: typically adds 2 weeks to the hiring process). U.S. Export Controls Disclaimer: In accordance with the U.S. Export Administration Regulations (15 CFR §743.13(b)), some roles at Intuitive Surgical may be subject to U.S. export controls for prospective employees who are nationals from countries currently on embargo or sanctions status.
Senior Research Scientist - Generative World Models For Autonomous Driving And Physical AI NvidiaSenior Research Scientist - Generative World Models For Autonomous Driving And Physical AISanta Clara, CAProven research excellence, evidenced by a strong publication record in leading AI and computer vision symposiums and periodicals (e.g., NeurIPS, CVPR, ICCV, ECCV, ICLR, ICML, TPAMI), along with experience translating research into impactful, real-world systems. This role requires outstanding software engineering abilities combined with deep expertise in innovative AI and simulation technologies, including generative world models, end-to-end driving, reasoning, and vision-language models.
Scientist 3, Neuro iPSC Disease Modeling Group / Stem Cell Neurobiology Genentech IncScientist 3, Neuro iPSC Disease Modeling Group / Stem Cell NeurobiologySouth San Francisco, CA$103,400–$192,000 / yearIn this role, you will lead the generation and application of complex iPSC-derived in vitro models and functional assays to understand the molecular mechanisms of neurodegenerative diseases, directly informing therapeutic interventions and pipeline advancement. Experience with automated liquid handlers, high-throughput screening workflows, or single-cell sequencing (scRNA-seq) analysis using R and/or Python.
Principal Architect Performance Analysis And Modeling D-MatrixPrincipal Architect Performance Analysis And ModelingSanta Clara, CAYour day-to-day work will include (1) analyzing the properties of emerging machine learning algorithms and workloads and identifying functional, performance implications (2) Creating analytical models to project performance on current and future generations of d-matrix hardware (3) proposing new HW/SW features to enable or accelerate these algorithms. Experience with developing analytical performance models, architecture simulators for performance analysis, Research background with publication record in top-tier architecture, or machine learning venues is a huge plus (such as ISCA, MICRO, ASPLOS, HPCA, DAC, MLSys etc.).
Model Policy, Frontier Cyber Risk OpenAIModel Policy, Frontier Cyber RiskSan Francisco, CAFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. Our open-plan offices have height-adjustable desks, conference rooms, phone booths, well-stocked kitchens full of snacks and drinks, three in-house prepared meals daily, a private outdoor space for working in the sun or socializing, nap rooms, private bike storage, and more.
Senior Financial Analyst, Gemini for Enterprise Model Serving Google LLCSenior Financial Analyst, Gemini for Enterprise Model ServingSunnyvale, CAWe are the Gemini for Enterprise product finance team that supports Google Cloud Platform (GCP)'s vast array of AI products including Gemini API, Gemini Enterprise Applications, Gemini Enterprise Solutions, Vertex Agent and Model Platforms, and other AI Solutions. We are an integral partner and strategic financial advisers to the Cloud leadership team, advising on how to invest resources to grow a strong and sustainable business.
Senior Applied Scientist, Delivery Foundation Model Amazon.com IncSenior Applied Scientist, Delivery Foundation ModelSanta Clara, CALead focused technical initiatives from conception through deployment, ensuring successful integration with production systems- Drive technical discussions within the team and and key stakeholders. In this role, you"ll combine highly technical work with scientific leadership, ensuring the team delivers robust solutions for dynamic real-world environments.
Principal Engineer, Model Development Platform Wayve Technologies LtdPrincipal Engineer, Model Development PlatformSunnyvale, CA$295,500–$335,300 / yearYou''ll lead by example, going deep across web applications, distributed compute, ML Ops, data pipelines, and optimization algorithms, and through architecture and mentorship you''ll enable teams to build platform capabilities that measurably accelerate model development and fleet learning. Experimentation & scheduling systems - Build systems that optimize how models are tested in simulation and on-road, using techniques like linear programming and heuristic optimization to balance hardware, safety, and research priorities while improving throughput and turnaround.
Strategy& - Strategy Consulting Business Model Reinvention - Manager PricewaterhouseCoopers LLPStrategy& - Strategy Consulting Business Model Reinvention - ManagerSan Francisco, CA$99,000–$232,000 / yearPwC does not intend to hire experienced or entry level job seekers who will need, now or in the future, PwC sponsorship through the H-1B lottery, except as set forth within the following policy: https://pwc.to/H-1B-Lottery-Policy . In this role, you will analyze client needs, provide consulting services across different strategic areas, and offer guidance to help clients develop and implement effective strategies that align with their business objectives.
Strategy& - Strategy Consulting Business Model Reinvention - Senior Associate PricewaterhouseCoopers LLPStrategy& - Strategy Consulting Business Model Reinvention - Senior AssociateSan Francisco, CA$77,000–$202,000 / yearIn this role at PwC, you will analyze client needs and provide consulting services across different strategic areas, offering guidance and support to help clients develop and implement effective strategies that align with their business objectives and drive growth. PwC does not intend to hire experienced or entry level job seekers who will need, now or in the future, PwC sponsorship through the H-1B lottery, except as set forth within the following policy: https://pwc.to/H-1B-Lottery-Policy .
Strategy& Strategy Consulting Business Model Reinvention - Director PricewaterhouseCoopers LLPStrategy& Strategy Consulting Business Model Reinvention - DirectorSan Francisco, CA$155,000–$410,000 / yearPwC does not intend to hire experienced or entry level job seekers who will need, now or in the future, PwC sponsorship through the H-1B lottery, except as set forth within the following policy: https://pwc.to/H-1B-Lottery-Policy . As a Director, you will set the strategic direction and lead business development efforts, making impactful decisions and overseeing multiple projects while maintaining executive-level client relations.
Research Intern - Video World Models (Research & ML Systems) 107752 Tencent LTDResearch Intern - Video World Models (Research & ML Systems) 107752Palo Alto, CA$80,168.40–$124,800 / yearWe are looking for "full-stack" hacker-researchers-visionary thinkers who are also elite engineers, capable of co-designing novel neural architectures and engineering the highly optimized infrastructure required to train them across large-scale computing cluster. Academic Excellence: Currently pursuing a PhD (or Master's degree with a truly exceptional research/engineering track record) in Computer Science, Machine Learning, Computer Architecture, or a related field.
Research Scientist - World-Action Foundation Model, Robotics Applied Intuition IncResearch Scientist - World-Action Foundation Model, RoboticsSunnyvale, CA$126,000–$423,000 / yearSupported by industry-leading tools and infra, researchers can access millions of miles of data from large fleets, and deploy methods they develop into various autonomous and robotic systems including self-driving cars/trucks, autonomous mining/construction machines, humanoid robots and dexterous hands. We're looking for someone who has: Strong research record in the fields of 3D vision, reconstruction and generation for robotics and autonomous systems, with publications in top-tier conferences or journals in the fields of computer vision, machine learning, and robotics.
Principal Architect, SoC and Systems Modelling NVIDIA CorpPrincipal Architect, SoC and Systems ModellingSanta Clara, CAIn this position, you will be working with other world-class architects on modeling, analysis and validation of chip & system architectures and features that advance the state of art in performance and efficiency. Develop tests, test plans, and testing infrastructure for new architectures/features and code coverage analysis and reporting.
Principal Architect, Soc And Systems Modelling NvidiaPrincipal Architect, Soc And Systems ModellingSanta Clara, CAIn this position, you will be working with other world-class architects on modeling, analysis and validation of chip & system architectures and features that advance the state of art in performance and efficiency. Develop tests, test plans, and testing infrastructure for new architectures/features and code coverage analysis and reporting.
Senior Scientist, Observing Systems & Inverse Modeling ReflectiveSenior Scientist, Observing Systems & Inverse ModelingSan Francisco, CaliforniaRemote$130,000–$180,000 / yearAs Reflective’s Senior Scientist, Observing Systems & Inverse Modeling, you will lead the design of observation strategies and inverse modeling workflows to determine what measurements are needed for SAI field experiments. Write scientific papers, concise memos, technical documentation, and public-facing summaries that make what has been learned, what remains uncertain, and how the results should inform experiment design clear to funders, policymakers, researchers, and the wider field.
Multimodal AI Model Optimization Research Engineer Tavus IncMultimodal AI Model Optimization Research EngineerSan Francisco, CAOur models power everything from text-to-video AI avatars to real-time conversational video experiences across industries like healthcare, recruiting, sales, and education. We achieve this through pioneering research in multimodal AI for modeling human-to-human communication (language, audio, and video), as well as generating audio-visual avatar behavior.
ASIC Firmware Engineer, Modeling OpenAIASIC Firmware Engineer, ModelingSan Francisco, CaliforniaFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. This role involves designing and developing drivers and functional models for a large array of HW components, writing high throughput and low latency firmware code, investigating bring-up and production issues.
Asic Firmware Engineer, Modeling OpenAIAsic Firmware Engineer, ModelingSan Francisco, CAFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. This role involves designing and developing drivers and functional models for a large array of HW components, writing high throughput and low latency firmware code, investigating bring-up and production issues.
Software Engineer, Model Hardware CoDesign Tesla IncSoftware Engineer, Model Hardware CoDesignPalo Alto, CADesign train and iterate on neural network architectures for autonomous driving and robotics with a focus on efficiency-aware model design architecture search distillation pruning quantization-aware training. Our team designs trains and deploys large-scale neural networks optimized for inference on compute-constrained edge devices CPU GPU custom AI ASIC.
Human Interactive Driving Intern - World Models Toyota Research InstituteHuman Interactive Driving Intern - World ModelsLos Altos, CA$45–$65 / hourDemonstrated experience with one or more of the following: World models (e.g., latent dynamics, diffusion-based models), Model-based RL or decision-making, 3D perception or sensor fusion, and Large-scale simulation for robotics or autonomous systems. Your work may focus on developing novel components of world models, improving decision-making through model-based RL, advancing 3D perception for dynamic scenes, or enhancing simulation-to-reality transfer.
Human Interactive Driving Intern – World Models Toyota Research InstituteHuman Interactive Driving Intern – World ModelsLos Altos, CA$45–$65 / hourDemonstrated experience with one or more of the following: World models (e.g., latent dynamics, diffusion-based models), Model-based RL or decision-making, 3D perception or sensor fusion, and Large-scale simulation for robotics or autonomous systems. Your work may focus on developing novel components of world models, improving decision-making through model-based RL, advancing 3D perception for dynamic scenes, or enhancing simulation-to-reality transfer.
Member of Technical Staff — Diffusion Model RadixArkMember of Technical Staff — Diffusion ModelPalo Alto, CaliforniaRadixArk is an infrastructure-first company built by engineers who've shipped production AI systems, created SGLang (30K+ GitHub stars, the fastest open LLM serving engine), and developed Miles (our large-scale RL framework). Our team has optimized kernels serving billions of tokens daily, designed distributed training systems coordinating 10,000+ GPUs, and contributed to infrastructure that powers leading AI companies and research labs.
Research Scientist - World-Action Foundation Model, Robotics Applied IntuitionResearch Scientist - World-Action Foundation Model, RoboticsSunnyvale, CaliforniaSupported by industry-leading tools and infra, researchers can access millions of miles of data from large fleets, and deploy methods they develop into various autonomous and robotic systems including self-driving cars/trucks, autonomous mining/construction machines, humanoid robots and dexterous hands. Applied Intuition is headquartered in Sunnyvale, California, with offices in Washington, D.C. San Diego; Ft. Walton Beach, Florida; Ann Arbor, Michigan; London; Stuttgart; Munich; Stockholm; Bangalore; Seoul; and Tokyo.
Research Scientist/Engineer, Neural Graphics and World Models Graduate (TikTok Engine & Tools) - 2027 Start (PhD) TikTok IncResearch Scientist/Engineer, Neural Graphics and World Models Graduate (TikTok Engine & Tools) - 2027 Start (PhD)San Jose, CADeep experience in at least one of the following AI-for-graphics areas: 3D modeling, asset generation, or 3D representations; animation or motion generation; simulation or physics-aware learning; rendering or neural rendering; world models or video prediction; or graphics and game-engine workflows. Familiarity with modern generative architectures and training methods such as variational autoencoders, latent tokenizers, diffusion or flow-matching models, diffusion transformers, autoregressive models, multimodal transformers, long-context modeling, pre-training, or post-training.
Research Scientist Intern (TikTok-Neural Graphics and World Models) - 2027 Start (PhD) TikTok IncResearch Scientist Intern (TikTok-Neural Graphics and World Models) - 2027 Start (PhD)San Jose, CADeep experience in at least one of the following AI-for-graphics areas: 3D modeling, asset generation, or 3D representations; animation or motion generation; simulation or physics-aware learning; rendering or neural rendering; world models or video prediction; or graphics and game-engine workflows. Familiarity with modern generative architectures and training methods such as variational autoencoders, latent tokenizers, diffusion or flow-matching models, diffusion transformers, autoregressive models, multimodal transformers, long-context modeling, pre-training, or post-training.
Quantitative Researcher - Model Scaling MSCI IncQuantitative Researcher - Model ScalingSan Francisco, CAThe Factors Research Platform and Governance team is responsible for building and modernizing the infrastructure, tooling, and frameworks that power MSCI's quantitative risk and factor model research. Hands-on experience with modern AI-assisted development tools and workflows, with a demonstrated ability to stay current with rapidly evolving AI capabilities and incorporate them into day-to-day work.
Backend Engineer, Models Meter, Inc.Backend Engineer, ModelsSan Francisco, CA$160,000–$230,000 / yearTo make this possible, we don't just need great models; we need infrastructure that gives those models clean, versioned, low-latency access to the right data, across training, evaluation, and deployment. As described on Meter.ai, we're building models in a closed-loop system that takes (as input) real-time telemetry, logs, and events on the network to autonomously troubleshoot, improve performance, and resolve issues.
World Model Research Scientist- Physical AI Kodiak Robotics, IncWorld Model Research Scientist- Physical AIMountain View, CA$180,000–$240,000 / yearShould the position require, and Kodiak determines that a candidate's residence, U.S. person status, and/or citizenship status necessitate an export license, bar the candidate from the position, or otherwise fall under national security-related restrictions, Kodiak will consider the candidate for alternative positions unaffected by such restrictions, under terms and conditions set forth at Kodiak's sole discretion, or, as an alternative, opt not to proceed with the candidate's application. We are looking for a research scientist to lead the design and development of world models capable of generating multi-sensor, multi-view, temporally coherent driving scenarios conditioned on actions, 3D scene context, and text.
ML Research Scientist - Quantum Accelerated Generative Models Sygaldry TechnologiesML Research Scientist - Quantum Accelerated Generative ModelsSan Francisco, CaliforniaYou'll identify where quantum approaches can provide genuine advantage in generative workflows—not incremental improvements, but structural speedups rooted in the mathematics of these models. Sygaldry AI servers combine multiple qubit types within a single, fault-tolerant architecture to deliver the combination of cost, scale, and speed necessary for advanced AI applications.
Senior Model-Based Systems Engineer Rondo Energy, Inc.Senior Model-Based Systems EngineerAlameda, CA$185,000–$210,000 / yearBuild and maintain automated workflows in an AI-powered systems engineering platform to connect requirements, interfaces, simulations, verification status, interface changes, and traceability into a live, continuously updated picture of program health. 7+ years in systems, integration, or verification engineering on complex multi-disciplinary, software-intensive systems (energy storage, industrial, aerospace, automotive, robotics, or similar); 3+ years in a model-based engineering environment.
Deals Services - Manager, Strategic Finance And Fp&A, Advanced Decision Modeling RSMDeals Services - Manager, Strategic Finance And Fp&A, Advanced Decision ModelingSan Francisco, CA$112,100–$225,500 / yearThe salary range (or starting rate for interns and associates) for this role represents numerous factors considered in the hiring decisions including, but not limited to, education, skills, work experience, certifications, location, etc. Develops and reviews complex, fully integrated financial models, including operating, cash flow, valuation, and transaction models, to support strategic decision-making.
Software Engineer (Model Evaluation & Benchmarking) SpreeAI CorpSoftware Engineer (Model Evaluation & Benchmarking)San Francisco, CAMultimodal generative systems require: benchmarking across visual realism, pose consistency, and identity preservation, automated regression detection across model checkpoints, scalable evaluation pipelines integrated into continuous deployment workflows. SPREEAI is a fast-growing, innovative AI company at the forefront of fashion and e-commerce, revolutionizing how consumers engage with fashion through lifelike photorealistic try-on technology and hyper-personalized shopping experiences.
Associate Director, Advanced Analytics, Patient Journey Insights and Predictive Modeling BeOne Medicines AGAssociate Director, Advanced Analytics, Patient Journey Insights and Predictive ModelingSan Carlos, CA$145,500–$195,500 / yearThis highly collaborative role partners with Commercial, Medical Affairs, Clinical Operations, Market Access, Medical Excellence, Clinical Development, IT, Data Engineering, Legal, Compliance, and external analytics partners to generate high-quality insights, build scalable predictive models, strengthen data and model governance, and embed analytics into strategic and operational decision-making. General Description: The Associate Director, Advanced Analytics - Patient Journey Insights and Predictive Modeling will develop and deliver advanced analytics, AI-enabled insights, and predictive modeling solutions that support data-driven decision-making across Commercial, Medical Affairs, and Clinical Operations.
Machine Learning Research Engineer, Model Evaluation WindBorne Systems IncMachine Learning Research Engineer, Model EvaluationRedwood City, CAWindBorne Systems is supercharging weather forecasts with a proprietary data source: a global constellation of next-generation smart weather balloons targeting critical atmospheric data. Evaluation strategy - Work with our Meteorology team to develop a rigorous, meteorologically valid strategy for comparing WeatherMesh with leading AI and physics-based models.
Software Engineer- Model Performance Systems BasetenSoftware Engineer- Model Performance SystemsSan Francisco, CaliforniaBy uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. You will not just be building the automated "speedometer and diagnostic" suite for our next-generation AI infrastructure; you will be defining the roadmap, driving key technical decisions, and taking full ownership of the future of this work.
VP, Bioassays, Cell Models, And Genomic Screening InsitroVP, Bioassays, Cell Models, And Genomic ScreeningSouth San Francisco, CA$290,000–$326,000 / yearLead innovation: Provide strategic, technical, and operational leadership in the development, validation, and implementation of disease-relevant cell models, genomic screening (phenotypic pooled optical screening and arrayed screening), and supporting assay development for drug discovery. In this leadership role, your primary responsibilities will be to: Drive insitro's discovery engine: Set the strategic vision for cell models, image based and genomic screens, as well as bioassay development and execution across all therapeutic areas, fueling our target and drug discovery ML engine.
NewSoftware Engineer - Model Products BasetenSoftware Engineer - Model ProductsSan Francisco, CaliforniaDesign, build, and operate the Model APIs surface with focus on advanced inference capabilities: structured outputs (JSON mode, grammar-constrained generation), tool/function calling and multi-modal serving. Profile and optimize TensorRT-LLM kernels, analyze CUDA kernel performance, implement custom CUDA operators, tune memory allocation patterns for maximum throughput and optimize communication patterns across multi-GPU setups.
Data Scientist, New Grad - Model Optimization quadric, IncData Scientist, New Grad - Model OptimizationBurlingame, CA$120,000–$160,000 / yearQuadric's co-optimized software and hardware is targeted to run neural network (NN) inference workloads in a wide variety of edge and endpoint devices, ranging from battery operated smart-sensor systems to high-performance automotive or autonomous vehicle systems. The actual base salary offered will depend on a number of factors, including the specific level of the role, years and depth of relevant experience, technical skills and competencies, the criticality of the role to the business, internal equity, and work location.
Machine Learning Research Engineer, Model Evaluation WindBorne SystemsMachine Learning Research Engineer, Model EvaluationPalo Alto, California$140,000–$240,000 / yearWindBorne Systems is supercharging weather forecasts with a proprietary data source: a global constellation of next-generation smart weather balloons targeting critical atmospheric data. Evaluation strategy — Work with our Meteorology team to develop a rigorous, meteorologically valid strategy for comparing WeatherMesh with leading AI and physics-based models.
Software Engineer - Model Performance BasetenSoftware Engineer - Model PerformanceSan Francisco, CaliforniaBy uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. Implement, refine, and productionize cutting-edge techniques (quantization, speculative decoding, kv cache reuse, chunked prefill and LoRA) for ML model inference and infrastructure.
Engineering Manager - Model Performance BasetenEngineering Manager - Model PerformanceSan Francisco, CaliforniaBy uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. Drive the development and deployment of large-scale optimization techniques for various ML models, especially large language models (LLMs).
Principal Research Engineer, Model Training & Post-Training Inflection AIPrincipal Research Engineer, Model Training & Post-TrainingPalo Alto, CA$400,000–$550,000 / yearLead training and post-training strategy, including supervised fine-tuning, RLHF, DPO, GRPO, RLAIF, reward modeling, preference optimization, tool-use fine-tuning, distillation, synthetic data, and related methods. Inflection's models are central to our product and platform strategy, and we are looking for a hands-on technical leader to own the model-improvement loop from data and training through evals, post-training, release criteria, and production feedback.
Sr. Software Engineer, Model Scaling, AI Infrastructure Tesla IncSr. Software Engineer, Model Scaling, AI InfrastructurePalo Alto, CA$176,000–$558,000 / yearYou will help build and optimize large-scale training systems running on thousands of GPUs while developing the tools, metrics, and workflows needed to make model scaling faster and more predictable. Drive improvements in model scaling efficiency, including larger models, longer context lengths, and higher-quality datasets.
Research Engineer - Audio & Speech Models Zyphra TechnologiesResearch Engineer - Audio & Speech ModelsSan Francisco, CAThe Role: As a Research Engineer - Audio & Speech Models, you will be a core contributor on Zyphra's Audio Team, building the next generation of open-source autoencoders, ASR, TTS, SSL, and speech-to-speech models. Qualifications / Additional Skills: Expertise and intuition for training models in the audio domain, including text-to-speech, ASR, speech-to-speech, speech-emotion-recognition, or other models.