Senior Applied Scientist, Efficient LLM Inference & Model Optimization Nebius Group NVSenior Applied Scientist, Efficient LLM Inference & Model OptimizationPalo Alto, CA$195,200–$262,200 / yearInvent, evaluate, and productionize methods for quantization, QAT, distillation, speculative decoding, KV-cache reuse, KV-cache compression, long-context inference, MoE routing, and model/runtime co-optimization. Build high-quality prototypes in PyTorch, Triton, CUDA-adjacent tooling, or inference-serving frameworks, then work with MLEs and platform engineers to productionize them.
Model Designer OpenAI LLCModel DesignerSan Francisco, CAFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity.
Foundation Models GenesisFoundation ModelsSan Carlos, CaliforniaDesign new generative simulation techniques to expand simulation data scale and diversity, training and evaluating generative models of 3D objects and environments, and language/code models to generate tasks and reward functions. Build machine learning models for robotics control end-to-end: data curation, careful evaluation, model architecture, training/inference stacks, rigorous experiments.
Engineering Manager, Model Inference AbridgeEngineering Manager, Model InferenceSan Francisco, CaliforniaYou’ll lead a high-performing team of AI inference engineers, partner closely with ML Research and the broader AI Platform, and ensure the systems underpinning every clinician interaction are operating at peak efficiency and reliability. The Inference team owns the end-to-end technical direction of how our models are served: from architecting low-latency, high-throughput infrastructure to pushing the frontier of LLM serving techniques.
Senior Manager, ML Occupancy Modeling, Autonomy RivianSenior Manager, ML Occupancy Modeling, AutonomyPalo Alto, California$265,000–$331,000 / yearFull timeRivian may use your Candidate Personal Data for the purposes of (i) tracking interactions with our recruiting system; (ii) carrying out, analyzing and improving our application and recruitment process, including assessing you and your application and conducting employment, background and reference checks; (iii) establishing an employment relationship or entering into an employment contract with you; (iv) complying with our legal, regulatory and corporate governance obligations; (v) recordkeeping; (vi) ensuring network and information security and preventing fraud; and (vii) as otherwise required or permitted by applicable law. Rivian may share your Candidate Personal Data with (i) internal personnel who have a need to know such information in order to perform their duties, including individuals on our People Team, Finance, Legal, and the team(s) with the position(s) for which you are applying; (ii) Rivian affiliates; and (iii) Rivian’s service providers, including providers of background checks, staffing services, and cloud services.
Staff Machine Learning Engineer, Diffusion, Generative Modeling And Inference SnapchatStaff Machine Learning Engineer, Diffusion, Generative Modeling And InferencePalo Alto, CA$229,000–$343,000 / yearOur team creates intuitive tools, platforms, and agentic systems that empower creators, developers, and internal teams to bring ideas to life, while advancing personalized, human-centric experiences across mobile, web, and wearable devices like Spectacles. 8+ years of post-Bachelor's machine learning or related experience; or a Master's degree in a technical field + 7+ years of post-grad ML or related experience; or a PhD in a related technical field + 4+ years of post-grad ML or related experience.
Research Scientist, RL for Autonomous Planning & World Modeling Waymo LLCResearch Scientist, RL for Autonomous Planning & World ModelingSan Francisco, CA$204,000–$259,000 / yearSince its start as the Google Self-Driving Car Project in 2009, Waymo has focused on building the Waymo Driver-The World''s Most Experienced Driver-to improve access to mobility while saving thousands of lives now lost to traffic crashes. Demonstration of original contributions to the field through high-impact publications (ArXiv, peer-reviewed conferences like NeurIPS/ICLR/CVPR), technical blog posts, or significant open-source contributions.
Senior / Principal ML Scientist, Foundation Models For Life Sciences LILASenior / Principal ML Scientist, Foundation Models For Life SciencesSan Francisco, CA$268,000–$384,000 / yearYou will shape the technical direction for how ML models are trained, evaluated, and deployed at scale, collaborate closely with AI scientists and experimental researchers to close the computational-experimental loop, and drive Lila's ML infrastructure toward the next generation of capabilities. Full-time U.S. employees receive a comprehensive benefits program including medical, dental, and vision coverage; employer-paid life and disability insurance; flexible time off with generous company wide holidays; paid parental leave; an educational assistance program; commuter benefits, including bike share memberships for office based employees; and a company subsidized lunch program.
Software Engineer, Models Meter, Inc.Software Engineer, ModelsSan Francisco, CAIn addition to your customers, network engineers, you'll partner closely with two research engineers who have deep ML backgrounds and a clear picture of what training data needs to look like. When a network engineer looks at a set of device stats and figures out it's upstream packet loss - not a hardware failure, not a misconfiguration, specifically upstream packet loss - that reasoning lives in their head.
Staff Software Engineer, Model Serving DataBricksStaff Software Engineer, Model ServingSan Francisco, CA$192,000–$260,000 / yearYou will design and build systems that enable high-throughput, low-latency inference across CPU and GPU workloads, influence architectural direction, and collaborate closely across platform, product, infrastructure, and research teams to deliver a world-class serving platform. Contribute directly to key components across the serving infrastructure - from model container builds and deployment workflows to runtime systems like routing, caching, observability, and intelligent autoscaling - ensuring smooth and efficient operations at scale.
Machine Learning: Multimodal Foundation Models The Bot CoMachine Learning: Multimodal Foundation ModelsSan Francisco, CAImprove Cross-Modal Reasoning: Research and implement methods to ensure the model doesnt just "associate" modalities but actually reasons through them (e.g., grounding visual physics in kinematic constraints).Own the Training Loop End-to-End: Design, run, debug, and iterate on large-scale training experiments; diagnosing failure modes, improving data mixtures, and tightening evaluation to drive measurable gains. Ship and Iterate on Real Systems: Integrate models into real robotic stacks, build on robot code to deploy your models, and optimize performance for edge inference.
Senior / Principal Scientist, Foundation Models for life Sciences Lila Sciences IncSenior / Principal Scientist, Foundation Models for life SciencesSan Francisco, CA$268,000–$384,000 / yearFull-time U.S. employees receive a comprehensive benefits program including medical, dental, and vision coverage; employer-paid life and disability insurance; flexible time off with generous company wide holidays; paid parental leave; an educational assistance program; commuter benefits, including bike share memberships for office based employees; and a company subsidized lunch program. LILA combines advanced AI models with proprietary AI Science Factory instruments into an operating system for science that executes the entire scientific method autonomously, accelerating discovery at unprecedented speed, scale, and impact across medicine, materials, and energy.
Principal Model Optimization Engineer Roblox CorpPrincipal Model Optimization EngineerSan Mateo, CA$295,250–$345,040 / yearEvery day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences- all created by our global community of developers and creators. A career at Roblox means you'll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone.
Staff Software Engineer, Foundation Model Inference DatabricksStaff Software Engineer, Foundation Model InferenceSan Francisco, CaliforniaBuild LLM infrastructure powering large-scale inference workloads for customers through partner models (OpenAI, Anthropic, Gemini) and self-hosted models (Qwen, GPT-OSS, Llama). More than 20,000 organizations worldwide — including adidas, AT&T, Bayer, Block, Mastercard, Rivian, Unilever, and 70% of the Fortune 500 — rely on the Databricks Data + AI Platform to build and scale data and AI apps, analytics and agents.
Senior Software Engineer, Model Serving DatabricksSenior Software Engineer, Model ServingSan Francisco, CaliforniaContribute directly to key components across the serving infrastructure — from model container builds and deployment workflows to runtime systems like routing, caching, observability, and intelligent autoscaling — ensuring smooth and efficient operations at scale. You will design and build systems that enable high-throughput, low-latency inference across CPU and GPU workloads, influence architectural direction, and collaborate closely across platform, product, infrastructure, and research teams to deliver a world-class serving platform.
ML Scientist I / II, Foundation Models for Life Sciences Lila Sciences IncML Scientist I / II, Foundation Models for Life SciencesSan Francisco, CA$176,000–$304,000 / yearFull-time U.S. employees receive a comprehensive benefits program including medical, dental, and vision coverage; employer-paid life and disability insurance; flexible time off with generous company wide holidays; paid parental leave; an educational assistance program; commuter benefits, including bike share memberships for office based employees; and a company subsidized lunch program. You will work on generative models spanning biological sequences, molecular structures, and multimodal experimental data, contributing to problem formulation, model design, training, evaluation, and integration into Lila''s closed-loop discovery engine.
Software Engineer - Voice Model TwitterSoftware Engineer - Voice ModelPalo Alto, CA$150,000–$450,000 / yearWork on pre-training and post-training of speech-language models, with targeted enhancements through supervised fine-tuning, reinforcement learning, and other techniques to ensure Grok Voice responses are accurate, factually grounded, natural and idiomatic in spoken style, conversational in tone, and fluent across multiple languages. Build and iterate a comprehensive evaluation framework covering objective metrics (accuracy, quality, latency, expressiveness), human preference studies, content factuality assessments, real-time interaction quality, and experimentation infrastructure to measure and improve performance.
Workload / Performance Model Lead SiFive IncWorkload / Performance Model LeadBerkeley, CA$231,444–$282,876 / yearSiFive's unrivaled compute platforms are continuing to enable leading technology companies around the world to innovate, optimize and deliver the most advanced solutions of tomorrow across every market segment of chip design, including artificial intelligence, machine learning, automotive, data center, mobile, and consumer. Any offer of employment for this position is also contingent on the Company verifying that you are a authorized for access to export-controlled technology under applicable export control laws or, if you are not already authorized, our ability to successfully obtain any necessary export license(s) or other approvals.
Software Engineer, Productivity - Model Performance OpenAISoftware Engineer, Productivity - Model PerformanceSan Francisco, CAFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity.
Engineering Manager, Model Flywheel OpenAIEngineering Manager, Model FlywheelSan Francisco, CaliforniaFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. The ChatGPT Model Flywheel team unified goal is to transform model advancements into great ChatGPT user experiences through reliable serving, rapid experimentation, safe deployment, and continuous improvement.