Engineering Manager - Model Performance BasetenEngineering Manager - Model PerformanceSan Francisco, CaliforniaBy uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. Drive the development and deployment of large-scale optimization techniques for various ML models, especially large language models (LLMs).
NewSoftware Engineer - Model Performance BasetenSoftware Engineer - Model PerformanceSan Francisco, CaliforniaBy uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. Implement, refine, and productionize cutting-edge techniques (quantization, speculative decoding, KV cache reuse, chunked prefill and LoRA) for ML model inference and infrastructure.
NewSoftware Engineer - Model Performance Systems BasetenSoftware Engineer - Model Performance SystemsSan Francisco, CaliforniaBy uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. You will not just be building the automated "speedometer and diagnostic" suite for our next-generation AI infrastructure; you will be defining the roadmap, driving key technical decisions, and taking full ownership of the future of this work.
NewTechnical Program Manager, Model Performance BasetenTechnical Program Manager, Model PerformanceSan Francisco, CaliforniaExperience program-managing model performance or inference optimization work - you understand how engines like vLLM, TensorRT-LLM, SGLang, or NVIDIA Dynamo fit into a production serving stack and can engage credibly with the engineers building on them. You won't inherit an existing program framework, you'll build one from the ground up: the planning structure, execution processes, metrics and the cross-functional alignment that a fast-growing organization needs.
NewSoftware Engineer - Model Products BasetenSoftware Engineer - Model ProductsSan Francisco, CaliforniaDesign, build, and operate the Model APIs surface with focus on advanced inference capabilities: structured outputs (JSON mode, grammar-constrained generation), tool/function calling and multi-modal serving. Profile and optimize TensorRT-LLM kernels, analyze CUDA kernel performance, implement custom CUDA operators, tune memory allocation patterns for maximum throughput and optimize communication patterns across multi-GPU setups.
Associate Product Manager, Hive Models HiveAssociate Product Manager, Hive ModelsSan Francisco, CA$90,000–$120,000 / yearAssociate Product Manager, Hive Models Role As an Associate Product Manager on our Hive Models team, you will work cross-functionally with stakeholders to help define product requirements and support implementation efforts, collaborating closely with our Machine Learning, Core Infrastructure, and Product teams. Support our Hive Models portfolio by learning the needs of the product development teams, and contributing to the creation of our deep learning models that provide human-like interpretation of video, image, audio, and text.
Member of Technical Staff (Model Behavior) PerplexityMember of Technical Staff (Model Behavior)San Francisco, CaliforniaContext Engineering: Design, test, and optimize the prompts, skills, tools, and memory that shape Perplexity responses across products, features, and use cases. We're hiring software engineers for the Model Behavior team to help shape how Perplexity’s AI products behave: the style of their responses, and the way they use tools, skills, and memory.
Research Engineer - Brain Computer Interface Models ZyphraResearch Engineer - Brain Computer Interface ModelsSan Francisco, CaliforniaAs a Research Engineer - Brain Computer Interface Models , you will be a core contributor to Zyphra’s BCI work, building the next generation of open-source EEG and brain–computer interface models. Ability to learn new domains quickly, orient oneself in the academic literature, and implement new ideas, creatively borrowing concepts from other modalities and applying them to EEG domain.
Research Engineer - Brain Computer Interface Models Zyphra TechnologiesResearch Engineer - Brain Computer Interface ModelsSan Francisco, CAThe Role: As a Research Engineer - Brain Computer Interface Models, you will be a core contributor to Zyphra's BCI work, building the next generation of open-source EEG and brain-computer interface models. Ability to learn new domains quickly, orient oneself in the academic literature, and implement new ideas, creatively borrowing concepts from other modalities and applying them to EEG domain.
Product Manager, Models HiveProduct Manager, ModelsSan Francisco, CA$120,000–$170,000 / yearAs a Product Manager on our Hive Models team, you will work cross-functionally with all stakeholders to define product requirements and see the implementation through to completion, leading development efforts between our Machine Learning, Core Infrastructure, and Product teams. Own our Hive Models portfolio, understanding the needs of the product development teams, and envisioning and executing the creation of our deep learning models that can provide human-like interpretation of video, image, audio and text.
Workload / Performance Model Lead SiFiveWorkload / Performance Model LeadBerkeley, CA$231,444–$282,876 / yearSiFive's unrivaled compute platforms are continuing to enable leading technology companies around the world to innovate, optimize and deliver the most advanced solutions of tomorrow across every market segment of chip design, including artificial intelligence, machine learning, automotive, data center, mobile, and consumer. Any offer of employment for this position is also contingent on the Company verifying that you are a authorized for access to export-controlled technology under applicable export control laws or, if you are not already authorized, our ability to successfully obtain any necessary export license(s) or other approvals.
Machinist / Model Maker, Master Keysight Technologies, Inc.Machinist / Model Maker, MasterSanta Rosa, California$94,640–$157,730 / yearKeysight Technologies is hiring a highly skilled, master-level Modelmaker (Machinist/Programmer) to support Swiss turning production on swing shift (2:00 PM–10:00 PM) at our Santa Rosa Precision Mesoscale Technology Center. Responsibilities: The Precision Meso-scale Technology Center (PMTC) located on Keysight’s campus in Santa Rosa, CA exists to provide mechanical solutions that enable the capabilities of Keysight’s industry leading electronic measurement equipment.
Sr. Scientist - Translational PBPK Modeling Genentech IncSr. Scientist - Translational PBPK ModelingSouth San Francisco, CAKey technical responsibilities are to strategize, plan, execute, and report PBPK modeling and simulations independently, as well as to present work at cross-functional teams, department meetings, senior management review committees, regulatory interactions, and scientific conferences. The Drug Metabolism and Pharmacokinetics (DMPK) group is dedicated to enabling the discovery, development and commercialization of safe and effective medicines by elucidating the absorption, distribution, metabolism, excretion and pharmacokinetic properties of small molecule drug candidates.
Senior Associate, Bess Modeling & Structuring Clearway Energy, Inc.Senior Associate, Bess Modeling & StructuringSan Francisco, CA$115,000–$150,000 / yearAlong with our public affiliate Clearway Energy, Inc., our portfolio comprises approximately 11.6 GW of gross generating capacity in 26 states, including 9.1 GW of wind, solar, and battery energy storage assets, and over 2.5 GW of conventional dispatchable power generation providing critical grid reliability services. Along with our public affiliate Clearway Energy, Inc., our portfolio comprises approximately 11.8 GW of gross generating capacity in 26 states, including 9.1 GW of wind, solar, and battery energy storage assets, and over 2.8 GW of flexible dispatchable power generation providing critical grid reliability services.
Software Engineer, Model Runtime OpenAI LLCSoftware Engineer, Model RuntimeSan Francisco, CAFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. You will work across model architecture, distributed systems, compilers, kernels, and silicon to design a production-grade runtime comparable in ambition to systems such as vLLM and SGLang, but customized and optimized for OpenAI's AI accelerator.
Research Engineer/Research Scientist, Personal AGI-Model Experience OpenAI LLCResearch Engineer/Research Scientist, Personal AGI-Model ExperienceSan Francisco, CAFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity.
Research Engineer / Research Scientist - Personal AGI, Personality and Model Behavior OpenAI LLCResearch Engineer / Research Scientist - Personal AGI, Personality and Model BehaviorSan Francisco, CAFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. Were looking for individuals with strong ML engineering skills and research experience, especially with novel and highly capable models, and in areas like reinforcement learning and reward modeling.
Program Officer, AI Model Safety Open rolesProgram Officer, AI Model SafetySan Francisco, CaliforniaBy AI model safety we mean the technical work of understanding and improving how models behave (testing and evaluating them independently, red-teaming them, researching how to interpret, align, and control them) so that increasingly capable systems can be trusted with increasingly consequential tasks. Proactively identify the most important projects and organizations that need to exist, and make them happen: scope priority projects, find or develop the right founders, and actively build new initiatives, seeding them if required.
Machine Learning: World Models The Bot CoMachine Learning: World ModelsSan Francisco, CAOwn the Training Loop End-to-End: Design, run, debug, and iterate on large-scale training experiments-diagnosing failure modes, improving data mixtures, and tightening evaluation to drive measurable gains. Architect Neural Simulators: Design and train spatiotemporal models that move beyond short clips toward coherent, long-form world simulations.
Member of Technical Staff (Model Behavior) Perplexity AI IncMember of Technical Staff (Model Behavior)San Francisco, CAContext Engineering: Design, test, and optimize the prompts, skills, tools, and memory that shape Perplexity responses across products, features, and use cases. Were hiring software engineers for the Model Behavior team to help shape how Perplexity's AI products behave: the style of their responses, and the way they use tools, skills, and memory.