NewSoftware Engineer, Model Runtime OpenAI LLCSoftware Engineer, Model RuntimeSan Francisco, CAFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. You will work across model architecture, distributed systems, compilers, kernels, and silicon to design a production-grade runtime comparable in ambition to systems such as vLLM and SGLang, but customized and optimized for OpenAI's AI accelerator.
Model Engineer - Member of Technical Staff MeterModel Engineer - Member of Technical StaffSan Francisco, CaliforniaEnd-to-end ownership : You'll ship models into production networks, collaborate with firmware and application teams sitting next to you, and rapidly iterate in the wild. Now, we’re assembling a founding core engineering team to build and train models that understand these systems, optimize operations, anticipate failures, and repair issues before humans even notice them.
Researcher, World Models MenloResearcher, World ModelsSan Francisco, CaliforniaConversant, ideally deep, in several of: SSL for visual and sensor representations; world models (JEPA, V-JEPA, I-JEPA, LeJEPA, MJEPA); generative and predictive architectures (diffusion, DiT, flow matching, VAEs); robotics ML (VLA, inverse dynamics, sim-to-real, optical flow); sensor fusion (vision, proprioception, force/torque, multi-modal encoders); PyTorch, JAX, and distributed training. Advance our self-supervised learning stack for visual and sensor representations, building on and extending the JEPA family (V-JEPA, I-JEPA, and related predictive-embedding approaches).
Software Engineer - Model Performance Systems BasetenSoftware Engineer - Model Performance SystemsSan Francisco, CaliforniaBy uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. You will not just be building the automated "speedometer and diagnostic" suite for our next-generation AI infrastructure; you will be defining the roadmap, driving key technical decisions, and taking full ownership of the future of this work.
Software Engineer - Model Products BasetenSoftware Engineer - Model ProductsSan Francisco, CaliforniaDesign, build, and operate the Model APIs surface with focus on advanced inference capabilities: structured outputs (JSON mode, grammar-constrained generation), tool/function calling and multi-modal serving. Profile and optimize TensorRT-LLM kernels, analyze CUDA kernel performance, implement custom CUDA operators, tune memory allocation patterns for maximum throughput and optimize communication patterns across multi-GPU setups.
Engineering Manager - Model Performance BasetenEngineering Manager - Model PerformanceSan Francisco, CaliforniaBy uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. Drive the development and deployment of large-scale optimization techniques for various ML models, especially large language models (LLMs).
Machine Learning Research Engineer, Model Evaluation WindBorne SystemsMachine Learning Research Engineer, Model EvaluationPalo Alto, CaliforniaWindBorne Systems is supercharging weather forecasts with a proprietary data source: a global constellation of next-generation smart weather balloons targeting critical atmospheric data. Evaluation strategy — Work with our Meteorology team to develop a rigorous, meteorologically valid strategy for comparing WeatherMesh with leading AI and physics-based models.
VP, Bioassays, Cell Models, And Genomic Screening InsitroVP, Bioassays, Cell Models, And Genomic ScreeningSouth San Francisco, CA$290,000–$326,000 / yearLead innovation: Provide strategic, technical, and operational leadership in the development, validation, and implementation of disease-relevant cell models, genomic screening (phenotypic pooled optical screening and arrayed screening), and supporting assay development for drug discovery. In this leadership role, your primary responsibilities will be to: Drive insitro's discovery engine: Set the strategic vision for cell models, image based and genomic screens, as well as bioassay development and execution across all therapeutic areas, fueling our target and drug discovery ML engine.
Software Engineer - Model Performance BasetenSoftware Engineer - Model PerformanceSan Francisco, CaliforniaBy uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. Implement, refine, and productionize cutting-edge techniques (quantization, speculative decoding, kv cache reuse, chunked prefill and LoRA) for ML model inference and infrastructure.
Machine Learning: World Models The Bot CompanyMachine Learning: World ModelsSan Francisco, CaliforniaOwn the Training Loop End-to-End: Design, run, debug, and iterate on large-scale training experiments—diagnosing failure modes, improving data mixtures, and tightening evaluation to drive measurable gains. Architect Neural Simulators: Design and train spatiotemporal models that move beyond short clips toward coherent, long-form world simulations.
Ai/Ml Engineer - Model Inference General MotorsAi/Ml Engineer - Model InferenceSunnyvale, CA$117,700–$221,400 / yearThe team sits at the intersection of machine learning, data infrastructure, and developer productivity, building systems that make it easier to search for important scenarios, prepare training-ready data, and support fast iteration across perception and evaluation workflows. We believe the next generation of autonomy and robotics depends not only on stronger models, but also on better infrastructure for turning massive volumes of multimodal data into reusable signals, searchable artifacts, and high-quality evaluation loops.
Principal Model Optimization Engineer RobloxPrincipal Model Optimization EngineerSan Mateo, CA$295,250–$345,040 / yearEvery day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone.
Senior AI Inference Engineer - Model Optimization & Deployment ZooxSenior AI Inference Engineer - Model Optimization & DeploymentFoster City, CAWe are looking for experts with hands-on experience in compressing, accelerating, and deploying complex models (LLMs, VLMs, or FMs) for power- and thermal-constrained vehicle SOCs. You will optimize the ML models, write custom CUDA kernels, and build highly concurrent inference code to ensure real-time, deterministic execution on edge devices.
Platform Engineer, Model Shaping Together AIPlatform Engineer, Model ShapingSan Francisco, CA$200,000–$290,000 / yearWe believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancements such as FlashAttention, RedPajama, SWARM Parallelism, and SpecExec.
Member of Technical Staff, Model Evaluation MirendilMember of Technical Staff, Model EvaluationSan Francisco, CaliforniaMirendil is a tech-first company focused on solving core bottlenecks that unlock step-change acceleration across science and technology. We are looking for a research engineer to build the evaluation infrastructure that tells us whether our models are getting better in ways we care about.
Research Scientist, Foundation Model (Video Generation) PikaResearch Scientist, Foundation Model (Video Generation)Palo Alto, CaliforniaDesign and prototype novel algorithms and architectures for high-fidelity, real-time multimodal synthesis and interaction across modalities. Advance state-of-the-art techniques in diffusion, autoregressive, and other generative models for large-scale pre-training and fine-tuning.
Associate, Quantitative Developer, Model Portfolio Solutions (MPS), Multi-Asset Strategies & Solutions (MASS) BlackRockAssociate, Quantitative Developer, Model Portfolio Solutions (MPS), Multi-Asset Strategies & Solutions (MASS)San Francisco, CaliforniaAs a Quantitative Developer on the Model Portfolio Solutions team, you will sit at the heart of BlackRock's innovation engine for quantitatively driven investing—designing, building, and scaling the signal implementations and analytical tools that researchers and portfolio managers rely on every day. You will translate cutting-edge quantitative research into production-grade signals, pioneer the integration of AI and agentic tooling into investment workflows, strengthen investment controls, and deliver scalable solutions that shape how a global, fast-growing systematic business invests.
Senior Product Manager, Model APIs & Developer Experience Together AISenior Product Manager, Model APIs & Developer ExperienceSan Francisco, California$200,000–$280,000 / yearWe believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. The interfaces around those models shape the entire customer experience: how easily a developer can migrate an application, how reliably an agent can use a model, and how a team runs large asynchronous workloads.
Research Engineering Manager - Model Training PerplexityResearch Engineering Manager - Model TrainingSan Francisco, CaliforniaExperience leading or managing research or engineering teams working on large-scale AI model development, including driving complex projects from idea to production. Lead a team of researchers and engineers focused on training SotA models for Perplexity-relevant use cases, leveraging the latest supervised and reinforcement learning techniques.
Technical Program Manager, Model Alignment And Deployment Character AITechnical Program Manager, Model Alignment And DeploymentRedwood City, CATogether, these groups are responsible for transforming powerful pretrained language models into intelligent, engaging, safely aligned, and highly scalable products-working across data, compute, algorithms, infrastructure, and user insights to improve model performance and ensure reliable delivery. Program ownership: Lead planning and execution of cross-functional programs spanning data collection, annotation pipelines, alignment workflows (RLHF, DPO, Constitutional AI), safety guardrails (adversarial testing, red-teaming), and model serving.