Senior Applied Scientist, Efficient LLM Inference & Model Optimization NebiusSenior Applied Scientist, Efficient LLM Inference & Model OptimizationPalo Alto, CaliforniaInvent, evaluate, and productionize methods for quantization, QAT, distillation, speculative decoding, KV-cache reuse, KV-cache compression, long-context inference, MoE routing, and model/runtime co-optimization. Build high-quality prototypes in PyTorch, Triton, CUDA-adjacent tooling, or inference-serving frameworks, then work with MLEs and platform engineers to productionize them.
Machine Learning Researcher / Engineer (Foundational Models) PathwayMachine Learning Researcher / Engineer (Foundational Models)Palo Alto, CARemotePathway is led by co-founder & CEO Zuzanna Stamirowska, a complexity scientist who created a team consisting of AI pioneers, including CTO Jan Chorowski who was the first person to apply Attention to speech and worked with Nobel laureate Geoff Hinton at Google Brain, as well as CSO Adrian Kosowski, a leading computer scientist and quantum physicist who obtained his PhD at the age of 20. The company is backed by leading investors and advisors, including Lukasz Kaiser, co-author of the Transformer (“the T” in ChatGPT) and a key researcher behind OpenAI’s reasoning models.
Manager, Multi-Modal Language Action Models ZooxManager, Multi-Modal Language Action ModelsFoster City, CAZoox is seeking an experienced Manager of Multi-Modal Language Action Models to lead a team focused on applying cutting-edge Large Multi-Modal Models (MLLMs) to solve concrete, offline autonomy problems. Proven track record of deploying ML/MLLM models to production or internal customers to solve complex, real-world problems, ideally within the autonomy or robotics domain.
Machine Learning Researcher / Engineer (Foundation Models) PathwayMachine Learning Researcher / Engineer (Foundation Models)Palo Alto, CARemotePathway is led by co-founder & CEO Zuzanna Stamirowska, a complexity scientist who created a team consisting of AI pioneers, including CTO Jan Chorowski who was the first person to apply Attention to speech and worked with Nobel laureate Geoff Hinton at Google Brain, as well as CSO Adrian Kosowski, a leading computer scientist and quantum physicist who obtained his PhD at the age of 20. The company is backed by leading investors and advisors, including Lukasz Kaiser, co-author of the Transformer (“the T” in ChatGPT) and a key researcher behind OpenAI’s reasoning models.
Software Engineer - Voice Model SpaceXAISoftware Engineer - Voice ModelPalo Alto, CA$150,000–$450,000 / yearWork on pre-training and post-training of speech-language models, with targeted enhancements through supervised fine-tuning, reinforcement learning, and other techniques to ensure Grok Voice responses are accurate, factually grounded, natural and idiomatic in spoken style, conversational in tone, and fluent across multiple languages. Build and iterate a comprehensive evaluation framework covering objective metrics (accuracy, quality, latency, expressiveness), human preference studies, content factuality assessments, real-time interaction quality, and experimentation infrastructure to measure and improve performance.
Backend Engineer, Models MeterBackend Engineer, ModelsSan Francisco, CaliforniaTo make this possible, we don’t just need great models; we need infrastructure that gives those models clean, versioned, low-latency access to the right data, across training, evaluation, and deployment. As described on Meter.ai , we’re building models in a closed-loop system that takes (as input) real-time telemetry, logs, and events on the network to autonomously troubleshoot, improve performance, and resolve issues.
Machine Learning Engineer – World Model Institute of Foundation ModelsMachine Learning Engineer – World ModelSunnyvale, CaliforniaStrategic and innovative problem-solving skills will be instrumental in establishing MBZUAI as a global hub for high-performance computing in deep learning, driving impactful discoveries that inspire the next generation of AI pioneers. You’ll build scalable, reliable, and observable cloud infrastructure, working closely with researchers to support data pipelines, experimentation, and evaluation workflows.
Research Engineer, Interactive World Models - New College Grad 2026 NvidiaResearch Engineer, Interactive World Models - New College Grad 2026Santa Clara, CAHelp advance the production-ready world model frontier by working with researchers on few-step distillation, causal or autoregressive generation, reward fine-tuning, action conditioning, and long-horizon spatiotemporal memory and consistency. Experience profiling or optimizing ML workloads using CUDA, Triton, TensorRT, torch.compile, or similar tools, including work on latency, throughput, quantization, streaming, state or cache management, or multi-GPU execution.
Senior Applied Research Scientist, Multimodal Foundation Models - Healthcare NVIDIA CorpSenior Applied Research Scientist, Multimodal Foundation Models - HealthcareSanta Clara, CAYou will collaborate with researchers, engineers, healthcare organizations, and industry partners to evaluate new ideas and translate successful research into software, models, and workflows that can be used by the broader healthcare ecosystem. What we need to see: PhD in Computer Science, Machine Learning, Biomedical Engineering, Computational Biology, Electrical Engineering, or a related quantitative field (or equivalent experience).
Research Engineer, Interactive World Models NvidiaResearch Engineer, Interactive World ModelsSanta Clara, CAContributions to an open-source ML project or developer platform, such as implementing model support, improving performance, building tests and benchmarks, fixing difficult issues, writing documentation, or helping users adopt the technology. Advance the production-ready world model frontier by working with researchers on few-step distillation, causal or autoregressive generation, reward fine-tuning, action conditioning, and long-horizon spatiotemporal memory and consistency.
Senior Research Engineer, Interactive World Models NvidiaSenior Research Engineer, Interactive World ModelsSanta Clara, CAAdvance the production-ready world model frontier by working with researchers on few-step distillation, causal or autoregressive generation, reward fine-tuning, action conditioning, and long-horizon spatiotemporal memory and consistency. What you'll be doing: Build and optimize the continuous autoregressive serving loop, including per-step control inputs, model and KV-cache state management, GPU inference, frame streaming, and model integrations to speed-of-light.
AI Systems Architect (Models & Hardware Co-Design) Velaura AI IncAI Systems Architect (Models & Hardware Co-Design)Sunnyvale, CAStaff / Principal)On-site - Full-timeBangalore / Santa Clara, CADVApplyDesign Verification EngineerOn-site - Full-timeSanta Clara, CA / Austin, Texas / Boston, MAApplyDesign Verification Engineer- AI Accelerator LeadOn-site - Full-timeSanta Clara, CAApplyDesign Verification Engineer- CPUOn-site - Full-timeSanta Clara, CAApplySenior Design Verification EngineerOn-site - Full-timeSanta Clara, CAApplySenior Formal Verification EngineerOn-site - Full-timeSanta Clara, CAEmulationApplySenior Emulation EngineerOn-site - Full-timeSanta Clara, CAPerformanceApplyPerformance Modeling Architect - AI SystemsOn-site - Full-timeDurham, North Carolina / Santa Clara, CA / Boston, MA / Austin, TexasRTLApplyRTL DesignerOn-site - Full-timeSanta Clara, CAApplyRTL LeadOn-site - Full-timeSanta Clara, CAApplyRTL Power EngineerOn-site - Full-timeSanta Clara, CAApplySenior RTL EngineerOn-site - Full-timeSanta Clara, CASoftwareCompiler & ToolchainApplyAccelerator Compiler and Tool Chain LeadOn-site - Full-timeSanta Clara, CAPlatform SoftwareApplyPlatform Software Lead - Physical AIOn-site - Full-timeSanta Clara, CAApplyPrincipal AI SoC Runtime Software ArchitectOn-site - Full-timeSanta Clara, CA. Manager)On-site - Full-timeBangalore / Santa Clara, CAApplyPhysical Design Engineer (Staff / Sr.
Model Maker Ursus, Inc.Model MakerFoster City, CA$52–$62.93 / hourYou will both bring and grow a prototyping mindset in addition to various capabilities including (but not limited to) metalwork (waterjet, shear, brake, mill, lathe, welding), composites, and additive manufacturing. Your support will include interfacing with both prototyping teammates and internal clients to support the rapid iteration for vehicle development.
Software Engineer - Model Performance Systems BasetenSoftware Engineer - Model Performance SystemsSan Francisco, CaliforniaBy uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. You will not just be building the automated "speedometer and diagnostic" suite for our next-generation AI infrastructure; you will be defining the roadmap, driving key technical decisions, and taking full ownership of the future of this work.
Software Engineer - Model Products BasetenSoftware Engineer - Model ProductsSan Francisco, CaliforniaDesign, build, and operate the Model APIs surface with focus on advanced inference capabilities: structured outputs (JSON mode, grammar-constrained generation), tool/function calling and multi-modal serving. Profile and optimize TensorRT-LLM kernels, analyze CUDA kernel performance, implement custom CUDA operators, tune memory allocation patterns for maximum throughput and optimize communication patterns across multi-GPU setups.
Engineering Manager - Model Performance BasetenEngineering Manager - Model PerformanceSan Francisco, CaliforniaBy uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. Drive the development and deployment of large-scale optimization techniques for various ML models, especially large language models (LLMs).
Machine Learning Research Engineer, Model Evaluation WindBorne SystemsMachine Learning Research Engineer, Model EvaluationPalo Alto, CaliforniaWindBorne Systems is supercharging weather forecasts with a proprietary data source: a global constellation of next-generation smart weather balloons targeting critical atmospheric data. Evaluation strategy — Work with our Meteorology team to develop a rigorous, meteorologically valid strategy for comparing WeatherMesh with leading AI and physics-based models.
VP, Bioassays, Cell Models, And Genomic Screening InsitroVP, Bioassays, Cell Models, And Genomic ScreeningSouth San Francisco, CA$290,000–$326,000 / yearLead innovation: Provide strategic, technical, and operational leadership in the development, validation, and implementation of disease-relevant cell models, genomic screening (phenotypic pooled optical screening and arrayed screening), and supporting assay development for drug discovery. In this leadership role, your primary responsibilities will be to: Drive insitro's discovery engine: Set the strategic vision for cell models, image based and genomic screens, as well as bioassay development and execution across all therapeutic areas, fueling our target and drug discovery ML engine.
Software Engineer - Model Performance BasetenSoftware Engineer - Model PerformanceSan Francisco, CaliforniaBy uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. Implement, refine, and productionize cutting-edge techniques (quantization, speculative decoding, kv cache reuse, chunked prefill and LoRA) for ML model inference and infrastructure.
Model Policy, Frontier Cyber Risk OpenAIModel Policy, Frontier Cyber RiskSan Francisco, CaliforniaFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. Our open-plan offices have height-adjustable desks, conference rooms, phone booths, well-stocked kitchens full of snacks and drinks, three in-house prepared meals daily, a private outdoor space for working in the sun or socializing, nap rooms, private bike storage, and more.