Staff Software Engineer, Model Serving DataBricksStaff Software Engineer, Model ServingSan Francisco, CA$192,000–$260,000 / yearYou will design and build systems that enable high-throughput, low-latency inference across CPU and GPU workloads, influence architectural direction, and collaborate closely across platform, product, infrastructure, and research teams to deliver a world-class serving platform. Contribute directly to key components across the serving infrastructure - from model container builds and deployment workflows to runtime systems like routing, caching, observability, and intelligent autoscaling - ensuring smooth and efficient operations at scale.
ML Engineer - Healthcare Data Curation & Model Workflows Stanford UniversityML Engineer - Healthcare Data Curation & Model WorkflowsStanford, CA$122,929–$145,389 / yearWORKING CONDITIONS: May be exposed to high voltage electricity, radiation or electromagnetic fields, lasers, noise > 80dB TWA, Allergens/Biohazards/Chemicals /Asbestos, confined spaces, working at heights 10 feet, temperature extremes, heavy metals, unusual work hours or routine overtime and/or inclement weather. Perform supervisory duties, including overseeing the work of technicians and other staff associated with the group/project, supervising the regular installation, maintenance, and operation of complex scientific or engineering projects, and training technicians, operators, and others working in particular scientific or engineering function area.
Research Scientist / Engineer – Foundation Model: Core Research LumaResearch Scientist / Engineer – Foundation Model: Core ResearchRedwood City, CaliforniaUnified Modeling & Efficiency Drive the core research that powers all of Luma's products — co-designing multimodal representations, advancing core algorithms for long-context training, and establishing rigorous scaling laws to predict performance across compute budgets. This role offers the chance to bridge frontier research with magical, shipped products like Dream Machine and Ray3, solving novel problems where no playbook exists.
Senior Machine Learning Engineer, Agentic Science/Generative Models, AI for Biology & Translation (AIBT) Genentech IncSenior Machine Learning Engineer, Agentic Science/Generative Models, AI for Biology & Translation (AIBT)South San Francisco, CA$168,100–$312,300 / yearThe new Computational Sciences Center of Excellence (CoE) is a strategic, unified group whose goal is to harness the transformative power of data and Artificial Intelligence (AI) to assist our scientists in both pRED and gRED to deliver more innovative and transformative medicines for patients worldwide. Roche's Research and Early Development organisations at Genentech (gRED) and Pharma (pRED) have demonstrated how these technologies accelerate R&D, leveraging data and novel computational models to drive impact.
Research Intern - World-Action Foundation Model, Robotics Applied Intuition IncResearch Intern - World-Action Foundation Model, RoboticsSunnyvale, CASupported by industry-leading tools and infra, researchers can access millions of miles of data from large fleets, and deploy methods they develop into various autonomous and robotic systems including self-driving cars/trucks, autonomous mining/construction machines, humanoid robots and dexterous hands. Applied Intuition is headquartered in Sunnyvale, California, with offices in Washington, D.C. San Diego; Ft. Walton Beach, Florida; Ann Arbor, Michigan; London; Stuttgart; Munich; Stockholm; Bangalore; Seoul; and Tokyo.
Model Policy, Frontier Cyber Risk OpenAI LLCModel Policy, Frontier Cyber RiskSan Francisco, CAFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. Our open-plan offices have height-adjustable desks, conference rooms, phone booths, well-stocked kitchens full of snacks and drinks, three in-house prepared meals daily, a private outdoor space for working in the sun or socializing, nap rooms, private bike storage, and more.
Principal Perception Engineer, Obstacle Foundation Models - Autonomous Vehicles NVIDIA CorpPrincipal Perception Engineer, Obstacle Foundation Models - Autonomous VehiclesSanta Clara, CAHands-on experience architecting and deploying DNN-based perception pipelines on embedded or real-time platforms, including optimization for latency, memory, and compute constraints, and experience with modern architectures such as CNNs and transformers, plus familiarity with techniques like large-scale pretraining, parameter-efficient fine-tuning (e.g., LoRA), or vision-language models (VLMs). Lead data strategy for perception: specify data and labeling requirements, prioritize data collection and annotation, and collaborate closely with data and ground-truth teams to maximize impact, including model-assisted workflows (e.g., active learning, auto-labeling, VLMs) and advanced model-in-the-loop tooling.
Research Engineer - Audio & Speech Models Zyphra TechnologiesResearch Engineer - Audio & Speech ModelsSan Francisco, CAThe Role: As a Research Engineer - Audio & Speech Models, you will be a core contributor on Zyphra's Audio Team, building the next generation of open-source autoencoders, ASR, TTS, SSL, and speech-to-speech models. Qualifications / Additional Skills: Expertise and intuition for training models in the audio domain, including text-to-speech, ASR, speech-to-speech, speech-emotion-recognition, or other models.
Machine Learning Research Engineer, Model Evaluation WindBorne SystemsMachine Learning Research Engineer, Model EvaluationPalo Alto, California$140,000–$240,000 / yearWindBorne Systems is supercharging weather forecasts with a proprietary data source: a global constellation of next-generation smart weather balloons targeting critical atmospheric data. Evaluation strategy — Work with our Meteorology team to develop a rigorous, meteorologically valid strategy for comparing WeatherMesh with leading AI and physics-based models.
NewSoftware Engineer - Model Performance Systems BasetenSoftware Engineer - Model Performance SystemsSan Francisco, CaliforniaBy uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. You will not just be building the automated "speedometer and diagnostic" suite for our next-generation AI infrastructure; you will be defining the roadmap, driving key technical decisions, and taking full ownership of the future of this work.
NewSoftware Engineer - Model Products BasetenSoftware Engineer - Model ProductsSan Francisco, CaliforniaDesign, build, and operate the Model APIs surface with focus on advanced inference capabilities: structured outputs (JSON mode, grammar-constrained generation), tool/function calling and multi-modal serving. Profile and optimize TensorRT-LLM kernels, analyze CUDA kernel performance, implement custom CUDA operators, tune memory allocation patterns for maximum throughput and optimize communication patterns across multi-GPU setups.
Staff ML Engineer, Search Ads Shopping Relevance Models Google LLCStaff ML Engineer, Search Ads Shopping Relevance ModelsMountain View, CAWe"re looking for engineers who bring fresh ideas from all areas, including information retrieval, distributed computing, large-scale system design, networking and data storage, security, artificial intelligence, natural language processing, UI design and mobile; the list goes on and is growing every day. Collaborate on user journey understanding, metric and label formulation, feature and model improvements, live traffic experiments, data analysis, tools and infrastructure, and more, to predict and improve user experience on search ads.
Principal Engineer, Model Development Platform Wayve Technologies LtdPrincipal Engineer, Model Development PlatformSunnyvale, CA$295,500–$335,300 / yearYou''ll lead by example, going deep across web applications, distributed compute, ML Ops, data pipelines, and optimization algorithms, and through architecture and mentorship you''ll enable teams to build platform capabilities that measurably accelerate model development and fleet learning. Experimentation & scheduling systems - Build systems that optimize how models are tested in simulation and on-road, using techniques like linear programming and heuristic optimization to balance hardware, safety, and research priorities while improving throughput and turnaround.
Senior Applied Scientist, Delivery Foundation Model Amazon.com IncSenior Applied Scientist, Delivery Foundation ModelSanta Clara, CALead focused technical initiatives from conception through deployment, ensuring successful integration with production systems- Drive technical discussions within the team and and key stakeholders. In this role, you"ll combine highly technical work with scientific leadership, ensuring the team delivers robust solutions for dynamic real-world environments.
Backend Engineer, Models MeterBackend Engineer, ModelsSan Francisco, CaliforniaTo make this possible, we don’t just need great models; we need infrastructure that gives those models clean, versioned, low-latency access to the right data, across training, evaluation, and deployment. As described on Meter.ai , we’re building models in a closed-loop system that takes (as input) real-time telemetry, logs, and events on the network to autonomously troubleshoot, improve performance, and resolve issues.
Software Engineer, Model Evaluation And Improvement BenchlingSoftware Engineer, Model Evaluation And ImprovementSan Francisco, CAYou'll work at the intersection of software engineering, biology, and frontier AI: finding tasks that are challenging for LLMs and valuable to scientists, designing evaluations that capture real scientific judgment, and building systems to create these tasks at scale. Over 200,000 scientists around the world trust Benchling to power their most important work, from academic labs to Sanofi, Moderna, and more than half of the world's top 50 biopharma.
Machine Learning Engineer – World Model Institute of Foundation ModelsMachine Learning Engineer – World ModelSunnyvale, CaliforniaStrategic and innovative problem-solving skills will be instrumental in establishing MBZUAI as a global hub for high-performance computing in deep learning, driving impactful discoveries that inspire the next generation of AI pioneers. You’ll build scalable, reliable, and observable cloud infrastructure, working closely with researchers to support data pipelines, experimentation, and evaluation workflows.
Senior Research Engineer, Interactive World Models NvidiaSenior Research Engineer, Interactive World ModelsSanta Clara, CAAdvance the production-ready world model frontier by working with researchers on few-step distillation, causal or autoregressive generation, reward fine-tuning, action conditioning, and long-horizon spatiotemporal memory and consistency. What you'll be doing: Build and optimize the continuous autoregressive serving loop, including per-step control inputs, model and KV-cache state management, GPU inference, frame streaming, and model integrations to speed-of-light.
Sr. Simulation & Model Training Engineer - Drone Autonomy (AI2531) SiMa Technologies IncSr. Simulation & Model Training Engineer - Drone Autonomy (AI2531)San Jose, CA$120,800–$193,300 / yearIn this role, you will bridge virtual testing and real-world flight by building realistic simulation environments and training models for reliable deployment on physical drone hardware. Simulation & Synthetic Data - Create physics-based virtual environments and sensor models to generate high-fidelity synthetic datasets for robust model training.
NewSenior Research Engineer, Foundation Model Training Infrastructure NvidiaSenior Research Engineer, Foundation Model Training InfrastructureSanta Clara, CAWays to stand out from the crowd: Master's or PhD's degree in Computer Science, Robotics, Engineering, or a related field; Demonstrated Tech Lead experience, coordinating a team of engineers and driving projects from conception to deployment; Strong experience at building large-scale LLM and multimodal LLM training infrastructure; Contributions to popular open-source AI frameworks or research publications in top-tier AI conferences, such as NeurIPS, ICRA, ICLR, CoRL. What we need to see: Bachelor's degree in Computer Science, Robotics, Engineering, or a related field; 10+ years of full-time industry experience in large-scale MLOps and AI infrastructure; Proven experience designing and optimizing distributed training systems with frameworks like PyTorch, JAX, or TensorFlow.