Senior Staff Software Engineer, AI Model Lifecycle Crusoe EnergySenior Staff Software Engineer, AI Model LifecycleSan Francisco, CA$237,600–$318,240 / yearAbout This Role: The Senior Staff Software Engineer for the AI Model Lifecycle team will play a crucial role in building a comprehensive managed platform for the entire application development lifecycle, with a specific focus on leveraging Machine Learning models, including Large Language Models (LLMs). We're looking for problem-solving, opportunity-finding teammates with a sense of urgency, who believe in the scale of our ambition and thrive on a path not fully paved - people who want to grow their careers alongside a team of experts across energy, manufacturing, data center construction, and cloud services.
Scientist /Senior Scientist, Multimodal & Relational Machine Learning Foundation Models Altos LabsScientist /Senior Scientist, Multimodal & Relational Machine Learning Foundation ModelsSan Francisco, CA$200,900–$257,500 / yearArchitect and implement novel hybrid models that integrate Large Language Models (LLMs) with Graph Neural Networks (GNNs) for multi-hop reasoning over biological knowledge graphs. Lead the design of efficient data loading strategies and distributed training recipes (e.g., FSDP, DeepSpeed) to train models across multiple GPU nodes.
Research Scientist - Vision Foundation Models Epsilon LabsResearch Scientist - Vision Foundation ModelsSan Francisco, CaliforniaDrive research and technical excellence through conference publications and technical blog posts, establishing best practices for training robust medical imaging models at scale. This role focuses on pretraining and scaling vision encoders for radiology diagnosis across X-ray, CT, and MRI, with a growing emphasis on 3D volumetric modeling.
Research Engineering Manager - Model Training PerplexityResearch Engineering Manager - Model TrainingSan Francisco, CaliforniaExperience leading or managing research or engineering teams working on large-scale AI model development, including driving complex projects from idea to production. Lead a team of researchers and engineers focused on training SotA models for Perplexity-relevant use cases, leveraging the latest supervised and reinforcement learning techniques.
Member of Technical Staff - Model Training SpaceXAIMember of Technical Staff - Model TrainingPalo Alto, CA$180,000–$600,000 / yearIf you previously trained models used by millions of people it's a big plus, but modeling experience is not required. SpaceXAI's mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge.
Tech Lead, Robotic AI Model Faraday FutureTech Lead, Robotic AI ModelFremont, CA$150,000–$180,000 / yearThe Company, together with its controlled subsidiaries, operates across multiple technology-driven areas, including AI electric vehicles, robotics, and its crypto business (AIXC), all under its upgraded Global EAI Industry Bridge Strategy, marking the beginning of a new chapter in AI mobility and Web3 integration. You will work across the full post-training lifecycle: curating demonstration data, fine-tuning vision-language-action (VLA) models or world models, training reinforcement learning policies in simulation, validating behaviors on real hardware, and optimizing models for on-robot inference.
Robotic AI Engineer/Applied Scientist - Foundation Models Maven RoboticsRobotic AI Engineer/Applied Scientist - Foundation ModelsSan Francisco, CaliforniaMaster Data Efficiency: Develop novel co-training strategies and efficient learning algorithms that leverage diverse data sources—from Internet-scale video to sparse, high-fidelity human interventions. You are not expected to be a master of every domain listed below; however, you must be able to justify world-class excellence in at least one core factor (e.g., model architecture, RL formulations, or high-scale data systems).
Senior Applied Scientist, Efficient LLM Inference & Model Optimization NebiusSenior Applied Scientist, Efficient LLM Inference & Model OptimizationPalo Alto, CaliforniaInvent, evaluate, and productionize methods for quantization, QAT, distillation, speculative decoding, KV-cache reuse, KV-cache compression, long-context inference, MoE routing, and model/runtime co-optimization. Build high-quality prototypes in PyTorch, Triton, CUDA-adjacent tooling, or inference-serving frameworks, then work with MLEs and platform engineers to productionize them.
Machine Learning Researcher / Engineer (Foundational Models) PathwayMachine Learning Researcher / Engineer (Foundational Models)Palo Alto, CARemotePathway is led by co-founder & CEO Zuzanna Stamirowska, a complexity scientist who created a team consisting of AI pioneers, including CTO Jan Chorowski who was the first person to apply Attention to speech and worked with Nobel laureate Geoff Hinton at Google Brain, as well as CSO Adrian Kosowski, a leading computer scientist and quantum physicist who obtained his PhD at the age of 20. The company is backed by leading investors and advisors, including Lukasz Kaiser, co-author of the Transformer (“the T” in ChatGPT) and a key researcher behind OpenAI’s reasoning models.
Manager, Multi-Modal Language Action Models ZooxManager, Multi-Modal Language Action ModelsFoster City, CAZoox is seeking an experienced Manager of Multi-Modal Language Action Models to lead a team focused on applying cutting-edge Large Multi-Modal Models (MLLMs) to solve concrete, offline autonomy problems. Proven track record of deploying ML/MLLM models to production or internal customers to solve complex, real-world problems, ideally within the autonomy or robotics domain.
Machine Learning Researcher / Engineer (Foundation Models) PathwayMachine Learning Researcher / Engineer (Foundation Models)Palo Alto, CARemotePathway is led by co-founder & CEO Zuzanna Stamirowska, a complexity scientist who created a team consisting of AI pioneers, including CTO Jan Chorowski who was the first person to apply Attention to speech and worked with Nobel laureate Geoff Hinton at Google Brain, as well as CSO Adrian Kosowski, a leading computer scientist and quantum physicist who obtained his PhD at the age of 20. The company is backed by leading investors and advisors, including Lukasz Kaiser, co-author of the Transformer (“the T” in ChatGPT) and a key researcher behind OpenAI’s reasoning models.
Software Engineer - Voice Model SpaceXAISoftware Engineer - Voice ModelPalo Alto, CA$150,000–$450,000 / yearWork on pre-training and post-training of speech-language models, with targeted enhancements through supervised fine-tuning, reinforcement learning, and other techniques to ensure Grok Voice responses are accurate, factually grounded, natural and idiomatic in spoken style, conversational in tone, and fluent across multiple languages. Build and iterate a comprehensive evaluation framework covering objective metrics (accuracy, quality, latency, expressiveness), human preference studies, content factuality assessments, real-time interaction quality, and experimentation infrastructure to measure and improve performance.
Backend Engineer, Models MeterBackend Engineer, ModelsSan Francisco, CaliforniaTo make this possible, we don’t just need great models; we need infrastructure that gives those models clean, versioned, low-latency access to the right data, across training, evaluation, and deployment. As described on Meter.ai , we’re building models in a closed-loop system that takes (as input) real-time telemetry, logs, and events on the network to autonomously troubleshoot, improve performance, and resolve issues.
Machine Learning Engineer – World Model Institute of Foundation ModelsMachine Learning Engineer – World ModelSunnyvale, CaliforniaStrategic and innovative problem-solving skills will be instrumental in establishing MBZUAI as a global hub for high-performance computing in deep learning, driving impactful discoveries that inspire the next generation of AI pioneers. You’ll build scalable, reliable, and observable cloud infrastructure, working closely with researchers to support data pipelines, experimentation, and evaluation workflows.
Research Engineer, Interactive World Models - New College Grad 2026 NvidiaResearch Engineer, Interactive World Models - New College Grad 2026Santa Clara, CAHelp advance the production-ready world model frontier by working with researchers on few-step distillation, causal or autoregressive generation, reward fine-tuning, action conditioning, and long-horizon spatiotemporal memory and consistency. Experience profiling or optimizing ML workloads using CUDA, Triton, TensorRT, torch.compile, or similar tools, including work on latency, throughput, quantization, streaming, state or cache management, or multi-GPU execution.
Senior Applied Research Scientist, Multimodal Foundation Models - Healthcare NVIDIA CorpSenior Applied Research Scientist, Multimodal Foundation Models - HealthcareSanta Clara, CAYou will collaborate with researchers, engineers, healthcare organizations, and industry partners to evaluate new ideas and translate successful research into software, models, and workflows that can be used by the broader healthcare ecosystem. What we need to see: PhD in Computer Science, Machine Learning, Biomedical Engineering, Computational Biology, Electrical Engineering, or a related quantitative field (or equivalent experience).
Research Engineer, Interactive World Models NvidiaResearch Engineer, Interactive World ModelsSanta Clara, CAContributions to an open-source ML project or developer platform, such as implementing model support, improving performance, building tests and benchmarks, fixing difficult issues, writing documentation, or helping users adopt the technology. Advance the production-ready world model frontier by working with researchers on few-step distillation, causal or autoregressive generation, reward fine-tuning, action conditioning, and long-horizon spatiotemporal memory and consistency.
Senior Research Engineer, Interactive World Models NvidiaSenior Research Engineer, Interactive World ModelsSanta Clara, CAAdvance the production-ready world model frontier by working with researchers on few-step distillation, causal or autoregressive generation, reward fine-tuning, action conditioning, and long-horizon spatiotemporal memory and consistency. What you'll be doing: Build and optimize the continuous autoregressive serving loop, including per-step control inputs, model and KV-cache state management, GPU inference, frame streaming, and model integrations to speed-of-light.
Senior Model-Based Systems Engineer Rondo Energy, Inc.Senior Model-Based Systems EngineerAlameda, CA$185,000–$210,000 / yearBuild and maintain automated workflows in an AI-powered systems engineering platform to connect requirements, interfaces, simulations, verification status, interface changes, and traceability into a live, continuously updated picture of program health. 7+ years in systems, integration, or verification engineering on complex multi-disciplinary, software-intensive systems (energy storage, industrial, aerospace, automotive, robotics, or similar); 3+ years in a model-based engineering environment.
ML Research Scientist - Quantum Accelerated Generative Models Sygaldry TechnologiesML Research Scientist - Quantum Accelerated Generative ModelsSan Francisco, CaliforniaYou'll identify where quantum approaches can provide genuine advantage in generative workflows—not incremental improvements, but structural speedups rooted in the mathematics of these models. Sygaldry AI servers combine multiple qubit types within a single, fault-tolerant architecture to deliver the combination of cost, scale, and speed necessary for advanced AI applications.