Program Officer, AI Model Safety Open rolesProgram Officer, AI Model SafetySan Francisco, CaliforniaBy AI model safety we mean the technical work of understanding and improving how models behave (testing and evaluating them independently, red-teaming them, researching how to interpret, align, and control them) so that increasingly capable systems can be trusted with increasingly consequential tasks. Proactively identify the most important projects and organizations that need to exist, and make them happen: scope priority projects, find or develop the right founders, and actively build new initiatives, seeding them if required.
Machine Learning: World Models The Bot CoMachine Learning: World ModelsSan Francisco, CAOwn the Training Loop End-to-End: Design, run, debug, and iterate on large-scale training experiments-diagnosing failure modes, improving data mixtures, and tightening evaluation to drive measurable gains. Architect Neural Simulators: Design and train spatiotemporal models that move beyond short clips toward coherent, long-form world simulations.
Member of Technical Staff (Model Behavior) Perplexity AI IncMember of Technical Staff (Model Behavior)San Francisco, CAContext Engineering: Design, test, and optimize the prompts, skills, tools, and memory that shape Perplexity responses across products, features, and use cases. Were hiring software engineers for the Model Behavior team to help shape how Perplexity's AI products behave: the style of their responses, and the way they use tools, skills, and memory.
Engineering Manager, AI Models Infrastructure IntercomEngineering Manager, AI Models InfrastructureDublin, CAFounded in 2011, Fin became one of the fastest growing companies and remains one of the largest private software companies in the world with nearly 30,000 global businesses using our products to transform their customer support. Fin can also be combined with our Helpdesk to become a complete solution called the Fin Customer Service Suite, which provides AI enhanced support for the more complex or high touch queries that require a human agent.
Staff Software Engineer, Foundational Model Serving DataBricksStaff Software Engineer, Foundational Model ServingSan Francisco, CA$192,000–$260,000 / yearWe're looking for engineers who have owned high scale operational sensitive systems like customer facing APIs, Edge Gateways, ML Inference, or similar services and have an interest in getting deep building LLM APIs and runtimes at scale. More than 20,000 organizations worldwide - including adidas, AT&T, Bayer, Block, Mastercard, Rivian, Unilever, and 70% of the Fortune 500 - rely on the Databricks Data + AI Platform to build and scale data and AI apps, analytics and agents.
Senior Product Manager, Model APIs & Developer Experience Together Computer IncSenior Product Manager, Model APIs & Developer ExperienceSan Francisco, CA$200,000–$280,000 / yearWe believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. The interfaces around those models shape the entire customer experience: how easily a developer can migrate an application, how reliably an agent can use a model, and how a team runs large asynchronous workloads.
Staff Software Engineer- Foundation Model Inference DataBricksStaff Software Engineer- Foundation Model InferenceSan Francisco, CA$190,000–$265,000 / yearThe impact you will have: Build LLM infrastructure powering large-scale inference workloads for customers through partner models (OpenAI, Anthropic, Gemini) and self-hosted models (Qwen, GPT-OSS, Llama). More than 20,000 organizations worldwide - including adidas, AT&T, Bayer, Block, Mastercard, Rivian, Unilever, and 70% of the Fortune 500 - rely on the Databricks Data + AI Platform to build and scale data and AI apps, analytics and agents.
Engineering Manager - Model Performance BasetenEngineering Manager - Model PerformanceSan Francisco, CaliforniaBy uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. Drive the development and deployment of large-scale optimization techniques for various ML models, especially large language models (LLMs).
NewTechnical Program Manager, Model Performance BasetenTechnical Program Manager, Model PerformanceSan Francisco, CaliforniaExperience program-managing model performance or inference optimization work - you understand how engines like vLLM, TensorRT-LLM, SGLang, or NVIDIA Dynamo fit into a production serving stack and can engage credibly with the engineers building on them. You won't inherit an existing program framework, you'll build one from the ground up: the planning structure, execution processes, metrics and the cross-functional alignment that a fast-growing organization needs.
NewSoftware Engineer - Model Performance BasetenSoftware Engineer - Model PerformanceSan Francisco, CaliforniaBy uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. Implement, refine, and productionize cutting-edge techniques (quantization, speculative decoding, KV cache reuse, chunked prefill and LoRA) for ML model inference and infrastructure.
NewSoftware Engineer - Model Performance Systems BasetenSoftware Engineer - Model Performance SystemsSan Francisco, CaliforniaBy uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. You will not just be building the automated "speedometer and diagnostic" suite for our next-generation AI infrastructure; you will be defining the roadmap, driving key technical decisions, and taking full ownership of the future of this work.
NewSoftware Engineer - Model Products BasetenSoftware Engineer - Model ProductsSan Francisco, CaliforniaDesign, build, and operate the Model APIs surface with focus on advanced inference capabilities: structured outputs (JSON mode, grammar-constrained generation), tool/function calling and multi-modal serving. Profile and optimize TensorRT-LLM kernels, analyze CUDA kernel performance, implement custom CUDA operators, tune memory allocation patterns for maximum throughput and optimize communication patterns across multi-GPU setups.
Software Engineer, Model Runtime OpenAI LLCSoftware Engineer, Model RuntimeSan Francisco, CAFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. You will work across model architecture, distributed systems, compilers, kernels, and silicon to design a production-grade runtime comparable in ambition to systems such as vLLM and SGLang, but customized and optimized for OpenAI's AI accelerator.
Research Engineer/Research Scientist, Personal AGI-Model Experience OpenAI LLCResearch Engineer/Research Scientist, Personal AGI-Model ExperienceSan Francisco, CAFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity.
Research Engineer / Research Scientist - Personal AGI, Personality and Model Behavior OpenAI LLCResearch Engineer / Research Scientist - Personal AGI, Personality and Model BehaviorSan Francisco, CAFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. Were looking for individuals with strong ML engineering skills and research experience, especially with novel and highly capable models, and in areas like reinforcement learning and reward modeling.
Software Engineer, Infrastructure Modelling Fluidstack LtdSoftware Engineer, Infrastructure ModellingSan Francisco, CAIf there is an error with your submission and you did not receive a confirmation email, please email careers@fluidstack.io with your resume/CV, the role youve applied for, and the date you submitted your application-- someone from our recruiting team will be in touch. Powerful AI will be the biggest lever for human choice weve ever built - but only if models are aligned with what humanity actually wants.
NewML Infra Engineer, Modeling Physical IntelligenceML Infra Engineer, ModelingSan Francisco, CaliforniaThe ML Infrastructure team supports and accelerates PI’s core modeling efforts by building the systems that make large-scale training reliable, reproducible, and fast. Own training/inference infrastructure: Design, implement, and maintain systems for large-scale model training, including scheduling, job management, checkpointing, and metrics/logging.
Lead Software Engineer, Model Serving Platform SciforiumLead Software Engineer, Model Serving PlatformSan Francisco, CaliforniaBacked by multi-million-dollar funding and direct sponsorship from AMD with hands-on support from AMD engineers the team is scaling rapidly to build the full stack powering frontier AI models and real-time applications. Experience with ML systems engineering, distributed GPU scheduling, open source inference engine like vLLM, Sglang, or TRT-LLM.
Machine Learning Infrastructure Engineer, Model Inference AbridgeMachine Learning Infrastructure Engineer, Model InferenceSan Francisco, CAAs an ML Infrastructure Engineer, Model Inference at Abridge, you'll play a pivotal role in building and optimizing the core inference infrastructure that powers our machine learning models. Powered by Linked Evidence and our purpose-built, auditable AI, we are the only company that maps AI-generated summaries to ground truth, helping providers quickly trust and verify the output.
Backend Engineer, Models MeterBackend Engineer, ModelsSan Francisco, CaliforniaTo make this possible, we don’t just need great models; we need infrastructure that gives those models clean, versioned, low-latency access to the right data, across training, evaluation, and deployment. As described on Meter.ai , we’re building models in a closed-loop system that takes (as input) real-time telemetry, logs, and events on the network to autonomously troubleshoot, improve performance, and resolve issues.