Engineering Manager, Model Flywheel OpenAI LLCEngineering Manager, Model FlywheelSan Francisco, CAFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. The ChatGPT Model Capabilities and Deployment team unified goal is to transform model advancements into great ChatGPT user experiences through reliable serving, rapid experimentation, safe deployment, and continuous improvement.
ML Systems Engineer, Large-Scale Model Training & RL Infrastructure Nebius Group NVML Systems Engineer, Large-Scale Model Training & RL InfrastructurePalo Alto, CA$195,200–$262,200 / yearIntegrate and extend frameworks such as Megatron-LM, DeepSpeed, PyTorch FSDP/DTensor, Ray, verl, slime, AReaL, OpenRLHF, or equivalent internal systems. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure.
Senior Software Engineer, Spatial Intelligence And Foundation Models NvidiaSenior Software Engineer, Spatial Intelligence And Foundation ModelsSanta Clara, CAYou will join a group of world-class robotics software and applied research engineers focused on geometric and semantic understanding, and reasoning for robots - building the perception systems that turn raw sensor data into actionable world understanding, shaping the future of physical AI! What you'll be doing: Design, implement, and deploy novel algorithms for spatial understanding, working on problems ranging from SLAM, structure-from-motion, optical flow, scene flow, and object reconstruction to training VLMs on a wide range of spatial reasoning skills.
Research Scientist - World-Action Foundation Model, Robotics Applied Intuition IncResearch Scientist - World-Action Foundation Model, RoboticsSunnyvale, CA$126,000–$423,000 / yearSupported by industry-leading tools and infra, researchers can access millions of miles of data from large fleets, and deploy methods they develop into various autonomous and robotic systems including self-driving cars/trucks, autonomous mining/construction machines, humanoid robots and dexterous hands. We're looking for someone who has: Strong research record in the fields of 3D vision, reconstruction and generation for robotics and autonomous systems, with publications in top-tier conferences or journals in the fields of computer vision, machine learning, and robotics.
Sr Machine Learning Manager, Data for Foundation Models - SIML Apple IncSr Machine Learning Manager, Data for Foundation Models - SIMLCupertino, CAM.S. or PhD in Computer Science or a related field such as Electrical Engineering, Robotics, Statistics, Applied Mathematics, or equivalent experience 3+ years of experience as an ML Manager, leading teams through modern computer vision and generative AI research Experience managing cross-functional projects, connecting initiatives to engineers interests, and helping engineers grow through mentorship and empathy 8+ years of experience training state-of-the-art generative or understanding models Proficiency in ML frameworks and distributed data processing pipelines as a hands on technical leader Excellent communication skills and an ability to drive clarity in ambiguous circumstancesPrior hands-on research experience working with data and building models Recent projects and experience with modern visual generative AI for images or videos Track record of strong industrial research, including delivery of complex projects or publications in top-tier venues. Typical projects involve leveraging billion-scale datasets, applying machine learning models to enrich data, and executing quality-focused data curation strategies.
Staff Applied Scientist, UI Control Models Apple IncStaff Applied Scientist, UI Control ModelsCupertino, CAMasters or PhD in Computer Science, Machine Learning Engineering or equivalent professional experience 7+ years of hands-on industry experience and a proven track record of shipping ML-powered products or features, ideally involving LLMs. You will lead research and engineering efforts in model post-training techniques that create powerful UI control, tool calling and compelling conversational models that run within the constraints of mobile hardware.
Backend Engineer, Models MeterBackend Engineer, ModelsSan Francisco, CaliforniaTo make this possible, we don’t just need great models; we need infrastructure that gives those models clean, versioned, low-latency access to the right data, across training, evaluation, and deployment. As described on Meter.ai , we’re building models in a closed-loop system that takes (as input) real-time telemetry, logs, and events on the network to autonomously troubleshoot, improve performance, and resolve issues.
Machine Learning Engineer – World Model Institute of Foundation ModelsMachine Learning Engineer – World ModelSunnyvale, CaliforniaStrategic and innovative problem-solving skills will be instrumental in establishing MBZUAI as a global hub for high-performance computing in deep learning, driving impactful discoveries that inspire the next generation of AI pioneers. You’ll build scalable, reliable, and observable cloud infrastructure, working closely with researchers to support data pipelines, experimentation, and evaluation workflows.
Senior Research Engineer, Interactive World Models NvidiaSenior Research Engineer, Interactive World ModelsSanta Clara, CAAdvance the production-ready world model frontier by working with researchers on few-step distillation, causal or autoregressive generation, reward fine-tuning, action conditioning, and long-horizon spatiotemporal memory and consistency. What you'll be doing: Build and optimize the continuous autoregressive serving loop, including per-step control inputs, model and KV-cache state management, GPU inference, frame streaming, and model integrations to speed-of-light.
Member of Technical Staff - Model Training X CorpMember of Technical Staff - Model TrainingPalo Alto, CA$180,000–$600,000 / yearIf you previously trained models used by millions of people it's a big plus, but modeling experience is not required. xAI's mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge.
ML Engineer - Healthcare Data Curation & Model Workflows Stanford UniversityML Engineer - Healthcare Data Curation & Model WorkflowsStanford, CA$122,929–$145,389 / yearWORKING CONDITIONS: May be exposed to high voltage electricity, radiation or electromagnetic fields, lasers, noise > 80dB TWA, Allergens/Biohazards/Chemicals /Asbestos, confined spaces, working at heights 10 feet, temperature extremes, heavy metals, unusual work hours or routine overtime and/or inclement weather. Perform supervisory duties, including overseeing the work of technicians and other staff associated with the group/project, supervising the regular installation, maintenance, and operation of complex scientific or engineering projects, and training technicians, operators, and others working in particular scientific or engineering function area.
Research Scientist / Engineer – Foundation Model: Core Research LumaResearch Scientist / Engineer – Foundation Model: Core ResearchRedwood City, CaliforniaUnified Modeling & Efficiency Drive the core research that powers all of Luma's products — co-designing multimodal representations, advancing core algorithms for long-context training, and establishing rigorous scaling laws to predict performance across compute budgets. This role offers the chance to bridge frontier research with magical, shipped products like Dream Machine and Ray3, solving novel problems where no playbook exists.
Senior Machine Learning Engineer, Agentic Science/Generative Models, AI for Biology & Translation (AIBT) Genentech IncSenior Machine Learning Engineer, Agentic Science/Generative Models, AI for Biology & Translation (AIBT)South San Francisco, CA$168,100–$312,300 / yearThe new Computational Sciences Center of Excellence (CoE) is a strategic, unified group whose goal is to harness the transformative power of data and Artificial Intelligence (AI) to assist our scientists in both pRED and gRED to deliver more innovative and transformative medicines for patients worldwide. Roche's Research and Early Development organisations at Genentech (gRED) and Pharma (pRED) have demonstrated how these technologies accelerate R&D, leveraging data and novel computational models to drive impact.
Research Intern - World-Action Foundation Model, Robotics Applied Intuition IncResearch Intern - World-Action Foundation Model, RoboticsSunnyvale, CASupported by industry-leading tools and infra, researchers can access millions of miles of data from large fleets, and deploy methods they develop into various autonomous and robotic systems including self-driving cars/trucks, autonomous mining/construction machines, humanoid robots and dexterous hands. Applied Intuition is headquartered in Sunnyvale, California, with offices in Washington, D.C. San Diego; Ft. Walton Beach, Florida; Ann Arbor, Michigan; London; Stuttgart; Munich; Stockholm; Bangalore; Seoul; and Tokyo.
Power Model Engineer SiFive IncPower Model EngineerSanta Clara, CA$178,848–$218,592 / yearSiFive's unrivaled compute platforms are continuing to enable leading technology companies around the world to innovate, optimize and deliver the most advanced solutions of tomorrow across every market segment of chip design, including artificial intelligence, machine learning, automotive, data center, mobile, and consumer. Any offer of employment for this position is also contingent on the Company verifying that you are a authorized for access to export-controlled technology under applicable export control laws or, if you are not already authorized, our ability to successfully obtain any necessary export license(s) or other approvals.
Software Engineer, Models MeterSoftware Engineer, ModelsSan Francisco, CaliforniaIn addition to your customers, network engineers, you’ll partner closely with two research engineers who have deep ML backgrounds and a clear picture of what training data needs to look like. When a network engineer looks at a set of device stats and figures out it’s upstream packet loss — not a hardware failure, not a misconfiguration, specifically upstream packet loss — that reasoning lives in their head.
Nvidia 2027 Internships: Ph.D. Research Large Language Models NvidiaNvidia 2027 Internships: Ph.D. Research Large Language ModelsSanta Clara, CAOur work in AI and digital twins is transforming the world's largest industries and profoundly impacting society - from gaming to robotics, self-driving cars to life-saving healthcare, climate change to virtual worlds where we can all connect and create. Depending on the internship, prior experience or knowledge requirements could include the following programming skills and technologies: Python, C++, CUDA, Deep Learning Framworks (PyTorch, Tensorflow, JAX, etc.).
Relational Foundation Model Engineer, Modern Data Stack NvidiaRelational Foundation Model Engineer, Modern Data StackSanta Clara, CAYou will partner with world-class researchers and engineers across the full machine learning lifecycle, from architecture exploration and large-scale training to post-training optimization and high-performance inference. What you'll be doing: Collaborate with researchers/engineers to enhance our Transformer and GNN-based models to operate seamlessly over any relational schema and heterogeneous graph.
Senior Applied Scientist, Efficient LLM Inference & Model Optimization Nebius Group NVSenior Applied Scientist, Efficient LLM Inference & Model OptimizationPalo Alto, CA$195,200–$262,200 / yearInvent, evaluate, and productionize methods for quantization, QAT, distillation, speculative decoding, KV-cache reuse, KV-cache compression, long-context inference, MoE routing, and model/runtime co-optimization. Build high-quality prototypes in PyTorch, Triton, CUDA-adjacent tooling, or inference-serving frameworks, then work with MLEs and platform engineers to productionize them.
Senior Software Engineer, NIM Model Customization NVIDIA CorpSenior Software Engineer, NIM Model CustomizationSanta Clara, CAExperience in utilizing model customization tools: LoRA, QLoRA, supervised fine-tuning (SFT), reinforcement learning (RL), preference optimization (e.g., DPO), and retrieval-augmented generation (RAG). We have some of the most forward-thinking and hardworking people in the world working for us and, due to unprecedented growth, our exclusive engineering teams are rapidly growing.