Research Scientist - Vision Foundation Models Epsilon LabsResearch Scientist - Vision Foundation ModelsSan Francisco, CaliforniaDrive research and technical excellence through conference publications and technical blog posts, establishing best practices for training robust medical imaging models at scale. This role focuses on pretraining and scaling vision encoders for radiology diagnosis across X-ray, CT, and MRI, with a growing emphasis on 3D volumetric modeling.
Scientist /Senior Scientist, Multimodal & Relational Machine Learning Foundation Models Altos LabsScientist /Senior Scientist, Multimodal & Relational Machine Learning Foundation ModelsSan Francisco, CA$200,900–$257,500 / yearArchitect and implement novel hybrid models that integrate Large Language Models (LLMs) with Graph Neural Networks (GNNs) for multi-hop reasoning over biological knowledge graphs. Lead the design of efficient data loading strategies and distributed training recipes (e.g., FSDP, DeepSpeed) to train models across multiple GPU nodes.
Senior Staff Software Engineer, AI Model Lifecycle Crusoe EnergySenior Staff Software Engineer, AI Model LifecycleSan Francisco, CA$237,600–$318,240 / yearAbout This Role: The Senior Staff Software Engineer for the AI Model Lifecycle team will play a crucial role in building a comprehensive managed platform for the entire application development lifecycle, with a specific focus on leveraging Machine Learning models, including Large Language Models (LLMs). We're looking for problem-solving, opportunity-finding teammates with a sense of urgency, who believe in the scale of our ambition and thrive on a path not fully paved - people who want to grow their careers alongside a team of experts across energy, manufacturing, data center construction, and cloud services.
Machine Learning: World Models The Bot CompanyMachine Learning: World ModelsSan Francisco, CaliforniaOwn the Training Loop End-to-End: Design, run, debug, and iterate on large-scale training experiments—diagnosing failure modes, improving data mixtures, and tightening evaluation to drive measurable gains. Architect Neural Simulators: Design and train spatiotemporal models that move beyond short clips toward coherent, long-form world simulations.
Model Designer OpenAIModel DesignerSan Francisco, CAFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity.
Research Program Manager - Adversarial Model Research OpenAIResearch Program Manager - Adversarial Model ResearchSan Francisco, CAFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. The Human Data team at OpenAI is responsible for identifying and mitigating risks in advanced AI systems by designing evaluations, surfacing vulnerabilities, and collaborating closely with researchers to strengthen model reliability and public trust.
Machine Learning Engineer, Model Evaluations (Speech LLM) - San Francisco PlaudMachine Learning Engineer, Model Evaluations (Speech LLM) - San FranciscoSan Francisco, CaliforniaWith a mission to amplify human intelligence, Plaud captures, structures, and compounds the intelligence generated in conversations — so humans can think better, decide faster, and execute with clarity. Can deeply partner with ML researchers to define exactly what "good" looks like for a Speech LLM, translating capabilities (like ASR robustness in noisy environments or TTS emotional steerability) into measurable benchmarks.
Solutions Architect - AI Model Specialist FriendliAISolutions Architect - AI Model SpecialistSan Francisco, CaliforniaOur infrastructure powers high-throughput, low-latency workloads for global organizations and integrates directly with Hugging Face, providing instant access to over 600,000 open-source models. You will work closely with our customers to integrate FriendliAI’s inference and agent frameworks into real-world products, enabling them to build and scale AI applications effectively.
Research Program Manager – Adversarial Model Research OpenAIResearch Program Manager – Adversarial Model ResearchSan Francisco, CaliforniaFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. The Human Data team at OpenAI is responsible for identifying and mitigating risks in advanced AI systems by designing evaluations, surfacing vulnerabilities, and collaborating closely with researchers to strengthen model reliability and public trust.
DevOps Engineer - AI Model Evaluator - AI Trainer MercorDevOps Engineer - AI Model Evaluator - AI TrainerSan Francisco, CaliforniaRemoteRegular use of AI coding agents such as Cursor, Claude Code, Codex, Windsurf, Gemini CLI, or similar tools. Use frontier AI coding agents to complete and evaluate complex infrastructure engineering tasks.
iOS Engineer - AI Model Evaluator MercoriOS Engineer - AI Model EvaluatorSan Francisco, CaliforniaRemoteRegular use of AI coding agents such as Cursor , Claude Code , Codex , Windsurf , Gemini CLI , or similar tools. Use frontier AI coding agents to complete and evaluate complex engineering tasks.
Data Engineer - AI Model Evaluation MercorData Engineer - AI Model EvaluationSan Francisco, CaliforniaRemoteReview model-generated implementations involving ETL pipelines , data warehouses , analytics platforms , and distributed data systems . Regular use of AI coding agents such as Cursor , Claude Code , Codex , Windsurf , Gemini CLI , or similar tools.
Principal Applied AI Researcher - Domain- Specific Models (Dublin, CA) Articul8Principal Applied AI Researcher - Domain- Specific Models (Dublin, CA)Dublin, CaliforniaMaintain hands-on research impact at the highest level — sustain a meaningful personal research contribution through technical work, publications, patents, and externally visible output, modeling what it means to be a world-class researcher who uses massively parallel agentic systems to achieve what was previously impossible. Mentor senior researchers and raise the ceiling on human potential — coach Staff and Senior researchers on designing agent-augmented research programs, raise the bar on technical judgment and experimental rigor, and shape hiring for researchers who are driven to redefine what's possible.
Software Engineer, Model Inference OpenAISoftware Engineer, Model InferenceSan Francisco, CaliforniaFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. We are looking for an engineer who wants to take the world's largest and most capable AI models and optimize them for use in a high-volume, low-latency, and high-availability production and research environment.
Member Of Technical Staff (Model Behavior) Perplexity AIMember Of Technical Staff (Model Behavior)San Francisco, CAContext Engineering: Design, test, and optimize the prompts, skills, tools, and memory that shape Perplexity responses across products, features, and use cases. We're hiring software engineers for the Model Behavior team to help shape how Perplexity's AI products behave: the style of their responses, and the way they use tools, skills, and memory.
Product Marketing Lead, API Models & Research OpenAIProduct Marketing Lead, API Models & ResearchSan Francisco, CaliforniaFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. Your scope will span frontier models, specialized models for cybersecurity and life sciences, multimodal capabilities, and the API portfolio that makes these systems accessible to developers and businesses.
Product Manager, Model Gateway Recruiting From ScratchProduct Manager, Model GatewaySan Francisco, New YorkYou'll define how engineering teams access frontier models, optimize reliability and cost, and establish the policies that allow dozens of product teams to build quickly while efficiently sharing AI infrastructure. Opportunity to own the AI infrastructure platform that powers every product across one of the world's fastest-growing AI companies, directly influencing engineering velocity, infrastructure efficiency, and the future of enterprise AI development.
Software Engineer- Model Performance Systems BaseTenSoftware Engineer- Model Performance SystemsSan Francisco, CABy uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. You will not just be building the automated "speedometer and diagnostic" suite for our next-generation AI infrastructure; you will be defining the roadmap, driving key technical decisions, and taking full ownership of the future of this work.
Product Manager, Claude Code Model Performance AnthropicProduct Manager, Claude Code Model PerformanceSan Francisco, CAAs a Product Manager on Claude Code's model performance team, you will drive model launches end-to-end, build evals that measure what matters, and partner directly with researchers and product engineers to translate model improvements into developer-facing outcomes. This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences.
Research Engineer, Production Model Post-Training AnthropicResearch Engineer, Production Model Post-TrainingSan Francisco, CAThis research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences. As a Research Engineer on our Post-Training team, you'll train our base models through the complete post-training stack to deliver the production Claude models that users interact with.