Machine Learning: World Models The Bot CompanyMachine Learning: World ModelsSan Francisco, CaliforniaOwn the Training Loop End-to-End: Design, run, debug, and iterate on large-scale training experiments—diagnosing failure modes, improving data mixtures, and tightening evaluation to drive measurable gains. Architect Neural Simulators: Design and train spatiotemporal models that move beyond short clips toward coherent, long-form world simulations.
Ai/Ml Engineer - Model Inference General MotorsAi/Ml Engineer - Model InferenceSunnyvale, CA$117,700–$221,400 / yearThe team sits at the intersection of machine learning, data infrastructure, and developer productivity, building systems that make it easier to search for important scenarios, prepare training-ready data, and support fast iteration across perception and evaluation workflows. We believe the next generation of autonomy and robotics depends not only on stronger models, but also on better infrastructure for turning massive volumes of multimodal data into reusable signals, searchable artifacts, and high-quality evaluation loops.
Director/Sr. Manager, AI Inference Model Scaling Cerebras SystemsDirector/Sr. Manager, AI Inference Model ScalingSunnyvale, CaliforniaWe build the compiler frontend, model transformation pipeline, graph optimization infrastructure, high-performance kernel enablement, and runtime integration that together make next-generation AI models execute with industry-leading performance. We collaborate closely with hardware architects, runtime engineers, cloud platform teams, AI researchers, and strategic customers to rapidly bring new model architectures into production.
Staff Software Engineer, Foundation Model Inference DataBricksStaff Software Engineer, Foundation Model InferenceSan Francisco, CA$190,000–$265,000 / yearThe impact you will have: Build LLM infrastructure powering large-scale inference workloads for customers through partner models (OpenAI, Anthropic, Gemini) and self-hosted models (Qwen, GPT-OSS, Llama). More than 20,000 organizations worldwide - including adidas, AT&T, Bayer, Block, Mastercard, Rivian, Unilever, and 70% of the Fortune 500 - rely on the Databricks Data + AI Platform to build and scale data and AI apps, analytics and agents.
Software Engineer - Voice Model TwitterSoftware Engineer - Voice ModelPalo Alto, CA$150,000–$450,000 / yearWork on pre-training and post-training of speech-language models, with targeted enhancements through supervised fine-tuning, reinforcement learning, and other techniques to ensure Grok Voice responses are accurate, factually grounded, natural and idiomatic in spoken style, conversational in tone, and fluent across multiple languages. Build and iterate a comprehensive evaluation framework covering objective metrics (accuracy, quality, latency, expressiveness), human preference studies, content factuality assessments, real-time interaction quality, and experimentation infrastructure to measure and improve performance.
Senior Software Engineer, Model Serving DataBricksSenior Software Engineer, Model ServingSan Francisco, CA$166,000–$225,000 / yearContribute directly to key components across the serving infrastructure - from model container builds and deployment workflows to runtime systems like routing, caching, observability, and intelligent autoscaling - ensuring smooth and efficient operations at scale. You will design and build systems that enable high-throughput, low-latency inference across CPU and GPU workloads, influence architectural direction, and collaborate closely across platform, product, infrastructure, and research teams to deliver a world-class serving platform.
NewSenior / Principal ML Scientist, Foundation Models for Life Sciences Lila SciencesSenior / Principal ML Scientist, Foundation Models for Life SciencesSan Francisco, California$268,000–$384,000 / yearYou will shape the technical direction for how ML models are trained, evaluated, and deployed at scale, collaborate closely with AI scientists and experimental researchers to close the computational–experimental loop, and drive Lila's ML infrastructure toward the next generation of capabilities. Full-time U.S. employees receive a comprehensive benefits program including medical, dental, and vision coverage; employer-paid life and disability insurance; flexible time off with generous company wide holidays; paid parental leave; an educational assistance program; commuter benefits, including bike share memberships for office based employees; and a company subsidized lunch program.
Full Stack Engineer, Scientific AI Models BenchlingFull Stack Engineer, Scientific AI ModelsSan Francisco, CaliforniaAlphaFold or Boltz2) predict structures, predict scientific properties, and generate new drug designs, acting as a design partner and a major time saver to scientists who are creating life-saving therapeutics. Projects you might work on include: adding new models as soon as they're published, improving model performance and scalability, and enabling scientists to automate their in-silico workflows by chaining models together into pipelines.
Software Engineer, Models MeterSoftware Engineer, ModelsSan Francisco, CaliforniaIn addition to your customers, network engineers, you’ll partner closely with two research engineers who have deep ML backgrounds and a clear picture of what training data needs to look like. When a network engineer looks at a set of device stats and figures out it’s upstream packet loss — not a hardware failure, not a misconfiguration, specifically upstream packet loss — that reasoning lives in their head.
NewSenior Research Scientist, Multi-Modal Language Models NvidiaSenior Research Scientist, Multi-Modal Language ModelsSanta Clara, CAOur team drives Nemotron Multi-modal technology and with your help, we will continue to drive our models to be state of the art open-source multi-modal models. We want to deliver models that work amazingly well in the real world right out of the box, and we also want to uplift the whole ecosystem of users of multi-modal LLMs.
Senior Deep Learning Engineer - Model Evaluation & AI Systems NvidiaSenior Deep Learning Engineer - Model Evaluation & AI SystemsSanta Clara, CAExperience acting as a technical bridge across teams or platforms (e.g., evaluation, training, or agent frameworks), combining architectural understanding with clear communication and influence. Ways to stand out from the crowd: Experience building or improving evaluation frameworks, benchmarks, or ML infrastructure used by other teams or external users.
Senior Radar Perception Engineer, Obstacle Foundation Models - Autonomous Vehicles NvidiaSenior Radar Perception Engineer, Obstacle Foundation Models - Autonomous VehiclesSanta Clara, CAEmbedded Optimization: Hands-on experience architecting and deploying DNN-based perception pipelines on embedded or real-time platforms, including optimization for latency, memory, and compute constraints, and familiarity with modern architectures (e.g., Transformers, BEV networks). NVIDIA GPUs run deep learning algorithms that simulate aspects of human intelligence, acting as the brain of computers, robots, and self-driving cars that can perceive and understand the world.
Senior Research Scientist, Multimodal Foundation Models And Robotics NvidiaSenior Research Scientist, Multimodal Foundation Models And RoboticsSanta Clara, CADeep understanding of robot kinematics, dynamics, and sensors; Ability to safely operate robot hardware, lab equipment, and tools; Knowledge of control methods, including PID, model predictive control, and whole-body control; Familiarity with physics simulation frameworks such as MuJoCo and Isaac Sim; Robot hardware design and hands-on building experience. Hands-on training experience and publications in at least one of the following topics: LLMs; Large vision-language models; Video generative models and diffusion algorithms; or Action-based transformers.
Senior Software Engineer, Spatial Intelligence And Foundation Models NvidiaSenior Software Engineer, Spatial Intelligence And Foundation ModelsSanta Clara, CAYou will join a group of world-class robotics software and applied research engineers focused on geometric and semantic understanding, and reasoning for robots - building the perception systems that turn raw sensor data into actionable world understanding, shaping the future of physical AI! What you'll be doing: Design, implement, and deploy novel algorithms for spatial understanding, working on problems ranging from SLAM, structure-from-motion, optical flow, scene flow, and object reconstruction to training VLMs on a wide range of spatial reasoning skills.
NewSenior Manager, Interactive World Model Platforms NvidiaSenior Manager, Interactive World Model PlatformsSanta Clara, CATechnical fluency in the ML primitives behind interactive world models, including diffusion or flow-matching models, autoregressive / causal video generation, self-forcing or causal-forcing style training, Gaussian splatting, NeRFs, and neural reconstruction. Ways to stand out from the crowd: Experience adopting Gaussian splats, NeRFs, neural reconstruction, neural shading, or other advanced rendering techniques for AV, robotics, simulation, synthetic data, or production rendering workflows.
Senior Applied Scientist, Efficient LLM Inference & Model Optimization Nebius Group NVSenior Applied Scientist, Efficient LLM Inference & Model OptimizationPalo Alto, CA$195,200–$262,200 / yearInvent, evaluate, and productionize methods for quantization, QAT, distillation, speculative decoding, KV-cache reuse, KV-cache compression, long-context inference, MoE routing, and model/runtime co-optimization. Build high-quality prototypes in PyTorch, Triton, CUDA-adjacent tooling, or inference-serving frameworks, then work with MLEs and platform engineers to productionize them.
Relational Foundation Model Engineer, Modern Data Stack NvidiaRelational Foundation Model Engineer, Modern Data StackSanta Clara, CAYou will partner with world-class researchers and engineers across the full machine learning lifecycle, from architecture exploration and large-scale training to post-training optimization and high-performance inference. What you'll be doing: Collaborate with researchers/engineers to enhance our Transformer and GNN-based models to operate seamlessly over any relational schema and heterogeneous graph.
Nvidia 2027 Internships: Ph.D. Research Large Language Models NvidiaNvidia 2027 Internships: Ph.D. Research Large Language ModelsSanta Clara, CAOur work in AI and digital twins is transforming the world's largest industries and profoundly impacting society - from gaming to robotics, self-driving cars to life-saving healthcare, climate change to virtual worlds where we can all connect and create. Depending on the internship, prior experience or knowledge requirements could include the following programming skills and technologies: Python, C++, CUDA, Deep Learning Framworks (PyTorch, Tensorflow, JAX, etc.).
Product Marketing Manager, Model API BasetenProduct Marketing Manager, Model APISan Francisco, CaliforniaThis role requires someone who is both technically savvy and strategic, with a proven track record of crafting compelling product narratives and building marketing assets that resonate with technical decision-makers. We are seeking an experienced Product Marketing Manager with a strong background in engaging developer audiences and delivering impactful go-to-market programs for native AI and enterprise companies.
Principal Model Optimization Engineer RobloxPrincipal Model Optimization EngineerSan Mateo, CA$295,250–$345,040 / yearEvery day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone.