Machine Learning Researcher / Engineer (Foundation Models) PathwayMachine Learning Researcher / Engineer (Foundation Models)Palo Alto, CARemotePathway is led by co-founder & CEO Zuzanna Stamirowska, a complexity scientist who created a team consisting of AI pioneers, including CTO Jan Chorowski who was the first person to apply Attention to speech and worked with Nobel laureate Geoff Hinton at Google Brain, as well as CSO Adrian Kosowski, a leading computer scientist and quantum physicist who obtained his PhD at the age of 20. The company is backed by leading investors and advisors, including Lukasz Kaiser, co-author of the Transformer (“the T” in ChatGPT) and a key researcher behind OpenAI’s reasoning models.
Machine Learning Researcher / Engineer (Foundational Models) PathwayMachine Learning Researcher / Engineer (Foundational Models)Palo Alto, CARemotePathway is led by co-founder & CEO Zuzanna Stamirowska, a complexity scientist who created a team consisting of AI pioneers, including CTO Jan Chorowski who was the first person to apply Attention to speech and worked with Nobel laureate Geoff Hinton at Google Brain, as well as CSO Adrian Kosowski, a leading computer scientist and quantum physicist who obtained his PhD at the age of 20. The company is backed by leading investors and advisors, including Lukasz Kaiser, co-author of the Transformer (“the T” in ChatGPT) and a key researcher behind OpenAI’s reasoning models.
Senior Applied Scientist, Efficient LLM Inference & Model Optimization NebiusSenior Applied Scientist, Efficient LLM Inference & Model OptimizationPalo Alto, California$195,200–$262,200 / yearInvent, evaluate, and productionize methods for quantization, QAT, distillation, speculative decoding, KV-cache reuse, KV-cache compression, long-context inference, MoE routing, and model/runtime co-optimization. Build high-quality prototypes in PyTorch, Triton, CUDA-adjacent tooling, or inference-serving frameworks, then work with MLEs and platform engineers to productionize them.
Robotic AI Engineer/Applied Scientist - Foundation Models Maven RoboticsRobotic AI Engineer/Applied Scientist - Foundation ModelsSan Francisco, CaliforniaMaster Data Efficiency: Develop novel co-training strategies and efficient learning algorithms that leverage diverse data sources—from Internet-scale video to sparse, high-fidelity human interventions. You are not expected to be a master of every domain listed below; however, you must be able to justify world-class excellence in at least one core factor (e.g., model architecture, RL formulations, or high-scale data systems).
Software Engineer - Voice Model SpaceXAISoftware Engineer - Voice ModelPalo Alto, CA$150,000–$450,000 / yearWork on pre-training and post-training of speech-language models, with targeted enhancements through supervised fine-tuning, reinforcement learning, and other techniques to ensure Grok Voice responses are accurate, factually grounded, natural and idiomatic in spoken style, conversational in tone, and fluent across multiple languages. Build and iterate a comprehensive evaluation framework covering objective metrics (accuracy, quality, latency, expressiveness), human preference studies, content factuality assessments, real-time interaction quality, and experimentation infrastructure to measure and improve performance.
Manager, Multi-Modal Language Action Models ZooxManager, Multi-Modal Language Action ModelsFoster City, CAZoox is seeking an experienced Manager of Multi-Modal Language Action Models to lead a team focused on applying cutting-edge Large Multi-Modal Models (MLLMs) to solve concrete, offline autonomy problems. Proven track record of deploying ML/MLLM models to production or internal customers to solve complex, real-world problems, ideally within the autonomy or robotics domain.
Senior Software Engineer, Spatial Intelligence And Foundation Models NvidiaSenior Software Engineer, Spatial Intelligence And Foundation ModelsSanta Clara, CAYou will join a group of world-class robotics software and applied research engineers focused on geometric and semantic understanding, and reasoning for robots - building the perception systems that turn raw sensor data into actionable world understanding, shaping the future of physical AI! What you'll be doing: Design, implement, and deploy novel algorithms for spatial understanding, working on problems ranging from SLAM, structure-from-motion, optical flow, scene flow, and object reconstruction to training VLMs on a wide range of spatial reasoning skills.
Software Engineer, Models MeterSoftware Engineer, ModelsSan Francisco, CaliforniaIn addition to your customers, network engineers, you’ll partner closely with two research engineers who have deep ML backgrounds and a clear picture of what training data needs to look like. When a network engineer looks at a set of device stats and figures out it’s upstream packet loss — not a hardware failure, not a misconfiguration, specifically upstream packet loss — that reasoning lives in their head.
Full Stack Engineer, Scientific AI Models BenchlingFull Stack Engineer, Scientific AI ModelsSan Francisco, CAAlphaFold or Boltz2) predict structures, predict scientific properties, and generate new drug designs, acting as a design partner and a major time saver to scientists who are creating life-saving therapeutics. Projects you might work on include: adding new models as soon as they're published, improving model performance and scalability, and enabling scientists to automate their in-silico workflows by chaining models together into pipelines.
Senior Radar Perception Engineer, Obstacle Foundation Models - Autonomous Vehicles NvidiaSenior Radar Perception Engineer, Obstacle Foundation Models - Autonomous VehiclesSanta Clara, CAEmbedded Optimization: Hands-on experience architecting and deploying DNN-based perception pipelines on embedded or real-time platforms, including optimization for latency, memory, and compute constraints, and familiarity with modern architectures (e.g., Transformers, BEV networks). NVIDIA GPUs run deep learning algorithms that simulate aspects of human intelligence, acting as the brain of computers, robots, and self-driving cars that can perceive and understand the world.
Senior Research Scientist, Multimodal Foundation Models And Robotics NvidiaSenior Research Scientist, Multimodal Foundation Models And RoboticsSanta Clara, CADeep understanding of robot kinematics, dynamics, and sensors; Ability to safely operate robot hardware, lab equipment, and tools; Knowledge of control methods, including PID, model predictive control, and whole-body control; Familiarity with physics simulation frameworks such as MuJoCo and Isaac Sim; Robot hardware design and hands-on building experience. Hands-on training experience and publications in at least one of the following topics: LLMs; Large vision-language models; Video generative models and diffusion algorithms; or Action-based transformers.
Senior Applied Scientist, Efficient LLM Inference & Model Optimization Nebius Group NVSenior Applied Scientist, Efficient LLM Inference & Model OptimizationPalo Alto, CA$195,200–$262,200 / yearInvent, evaluate, and productionize methods for quantization, QAT, distillation, speculative decoding, KV-cache reuse, KV-cache compression, long-context inference, MoE routing, and model/runtime co-optimization. Build high-quality prototypes in PyTorch, Triton, CUDA-adjacent tooling, or inference-serving frameworks, then work with MLEs and platform engineers to productionize them.
Relational Foundation Model Engineer, Modern Data Stack NvidiaRelational Foundation Model Engineer, Modern Data StackSanta Clara, CAYou will partner with world-class researchers and engineers across the full machine learning lifecycle, from architecture exploration and large-scale training to post-training optimization and high-performance inference. What you'll be doing: Collaborate with researchers/engineers to enhance our Transformer and GNN-based models to operate seamlessly over any relational schema and heterogeneous graph.
Nvidia 2027 Internships: Ph.D. Research Large Language Models NvidiaNvidia 2027 Internships: Ph.D. Research Large Language ModelsSanta Clara, CAOur work in AI and digital twins is transforming the world's largest industries and profoundly impacting society - from gaming to robotics, self-driving cars to life-saving healthcare, climate change to virtual worlds where we can all connect and create. Depending on the internship, prior experience or knowledge requirements could include the following programming skills and technologies: Python, C++, CUDA, Deep Learning Framworks (PyTorch, Tensorflow, JAX, etc.).
NewSenior Research Scientist, Multi-Modal Language Models NvidiaSenior Research Scientist, Multi-Modal Language ModelsSanta Clara, CAOur team drives Nemotron Multi-modal technology and with your help, we will continue to drive our models to be state of the art open-source multi-modal models. We want to deliver models that work amazingly well in the real world right out of the box, and we also want to uplift the whole ecosystem of users of multi-modal LLMs.
NewSenior Manager, Interactive World Model Platforms NvidiaSenior Manager, Interactive World Model PlatformsSanta Clara, CATechnical fluency in the ML primitives behind interactive world models, including diffusion or flow-matching models, autoregressive / causal video generation, self-forcing or causal-forcing style training, Gaussian splatting, NeRFs, and neural reconstruction. Ways to stand out from the crowd: Experience adopting Gaussian splats, NeRFs, neural reconstruction, neural shading, or other advanced rendering techniques for AV, robotics, simulation, synthetic data, or production rendering workflows.
Senior / Principal ML Scientist, Foundation Models for Life Sciences Lila SciencesSenior / Principal ML Scientist, Foundation Models for Life SciencesSan Francisco, CaliforniaYou will shape the technical direction for how ML models are trained, evaluated, and deployed at scale, collaborate closely with AI scientists and experimental researchers to close the computational–experimental loop, and drive Lila's ML infrastructure toward the next generation of capabilities. Full-time U.S. employees receive a comprehensive benefits program including medical, dental, and vision coverage; employer-paid life and disability insurance; flexible time off with generous company wide holidays; paid parental leave; an educational assistance program; commuter benefits, including bike share memberships for office based employees; and a company subsidized lunch program.
Senior Deep Learning Engineer - Model Evaluation & AI Systems NvidiaSenior Deep Learning Engineer - Model Evaluation & AI SystemsSanta Clara, CAExperience acting as a technical bridge across teams or platforms (e.g., evaluation, training, or agent frameworks), combining architectural understanding with clear communication and influence. Ways to stand out from the crowd: Experience building or improving evaluation frameworks, benchmarks, or ML infrastructure used by other teams or external users.
Senior Software Engineer - Model Performance InferenceSenior Software Engineer - Model PerformanceSan Francisco, CaliforniaYour work spans from implementing known optimization techniques to experimenting with novel approaches, always with the goal of serving models faster and cheaper at scale. If you love squeezing every last drop of performance out of GPUs, diving deep into CUDA kernels, and turning optimization techniques into production systems, we'd love to meet you.
Machine Learning: Multimodal Foundation Models The Bot CompanyMachine Learning: Multimodal Foundation ModelsSan Francisco, CaliforniaImprove Cross-Modal Reasoning: Research and implement methods to ensure the model doesn't just "associate" modalities but actually reasons through them (e.g., grounding visual physics in kinematic constraints). Ship and Iterate on Real Systems: Integrate models into real robotic stacks, build on robot code to deploy your models, and optimize performance for edge inference.