Member of Technical Staff, Model Evaluation MirendilMember of Technical Staff, Model EvaluationSan Francisco, CaliforniaMirendil is a tech-first company focused on solving core bottlenecks that unlock step-change acceleration across science and technology. We are looking for a research engineer to build the evaluation infrastructure that tells us whether our models are getting better in ways we care about.
Research Scientist, Foundation Model (Video Generation) PikaResearch Scientist, Foundation Model (Video Generation)Palo Alto, CaliforniaDesign and prototype novel algorithms and architectures for high-fidelity, real-time multimodal synthesis and interaction across modalities. Advance state-of-the-art techniques in diffusion, autoregressive, and other generative models for large-scale pre-training and fine-tuning.
Associate, Quantitative Developer, Model Portfolio Solutions (MPS), Multi-Asset Strategies & Solutions (MASS) BlackRockAssociate, Quantitative Developer, Model Portfolio Solutions (MPS), Multi-Asset Strategies & Solutions (MASS)San Francisco, CaliforniaAs a Quantitative Developer on the Model Portfolio Solutions team, you will sit at the heart of BlackRock's innovation engine for quantitatively driven investing—designing, building, and scaling the signal implementations and analytical tools that researchers and portfolio managers rely on every day. You will translate cutting-edge quantitative research into production-grade signals, pioneer the integration of AI and agentic tooling into investment workflows, strengthen investment controls, and deliver scalable solutions that shape how a global, fast-growing systematic business invests.
Senior Product Manager, Model APIs & Developer Experience Together AISenior Product Manager, Model APIs & Developer ExperienceSan Francisco, California$200,000–$280,000 / yearWe believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. The interfaces around those models shape the entire customer experience: how easily a developer can migrate an application, how reliably an agent can use a model, and how a team runs large asynchronous workloads.
Senior Product Manager, Model Apis & Developer Experience Together AISenior Product Manager, Model Apis & Developer ExperienceSan Francisco, CA$200,000–$280,000 / yearWe believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. The interfaces around those models shape the entire customer experience: how easily a developer can migrate an application, how reliably an agent can use a model, and how a team runs large asynchronous workloads.
Embodied AI Researcher (VLA Models) Hyphen Connect LimitedEmbodied AI Researcher (VLA Models)San Francisco, CaliforniaThe successful candidate will contribute to cutting-edge research, designing models and experiments that advance the field of robotics through deep learning and zero-shot generalization. Strong programming skills in Python and familiarity with machine learning frameworks such as TensorFlow or PyTorch.
Account Sales Executive- San Francisco In-Vivo Laboratory Animal Models and Laboratory Consumables RPM ReSearchAccount Sales Executive- San Francisco In-Vivo Laboratory Animal Models and Laboratory ConsumablesSan Francisco, CaliforniaSearching for a professional, forward-thinking, enthusiastic highly motivated sales performer to manage and grow an established marquee territory in the Greater San Francisco Region selling technical models into Contract Research, Biopharma, Biotech, and Academic Accounts to further preclinical research and development. We are looking for a successful sales history in the life science marketplace working with large Academic, Biotechnology and Biopharmaceutical accounts.
Engineering Manager, Model Infrastructure HarveyEngineering Manager, Model InfrastructureSan Francisco, California$260,000–$340,000 / yearExperience working with multiple model providers such as OpenAI, Anthropic, Azure OpenAI, Fireworks, Baseten, or open-source model ecosystems. Drive the evolution of our Unified Model Controller (UMC) and Model Selector platform to automatically detect degraded models and intelligently route traffic based on health, latency, quality, compliance, and cost.
NewSoftware Engineer, Model Evaluation and Improvement BenchlingSoftware Engineer, Model Evaluation and ImprovementSan Francisco, CaliforniaYou’ll work at the intersection of software engineering, biology, and frontier AI: finding tasks that are challenging for LLMs and valuable to scientists, designing evaluations that capture real scientific judgment, and building systems to create these tasks at scale. Over 200,000 scientists around the world trust Benchling to power their most important work, from academic labs to Sanofi, Moderna, and more than half of the world's top 50 biopharma.
Product Manager, Growth, Scientific AI Models BenchlingProduct Manager, Growth, Scientific AI ModelsSan Francisco, CAOwn the product roadmap for our scientific model platform - defining and prioritizing the models, features, and capabilities that let scientists run inference fast, reliably, and cost-effectively across a diverse set of scientific models. AlphaFold or Boltz2) predict structures, predict scientific properties, and generate new drug designs, acting as a design partner and a major time saver to scientists who are creating life-saving therapeutics.
NewSoftware Engineer, Model Evaluation And Improvement BenchlingSoftware Engineer, Model Evaluation And ImprovementSan Francisco, CAYou'll work at the intersection of software engineering, biology, and frontier AI: finding tasks that are challenging for LLMs and valuable to scientists, designing evaluations that capture real scientific judgment, and building systems to create these tasks at scale. Over 200,000 scientists around the world trust Benchling to power their most important work, from academic labs to Sanofi, Moderna, and more than half of the world's top 50 biopharma.
Research Engineer, Production Model Post-Training AnthropicResearch Engineer, Production Model Post-TrainingSan Francisco, CAThis research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences. As a Research Engineer on our Post-Training team, you'll train our base models through the complete post-training stack to deliver the production Claude models that users interact with.
Research Product Manager, Model Behaviors AnthropicResearch Product Manager, Model BehaviorsSan Francisco, CAThis research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences. We offer competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space in which to collaborate with colleagues.
Product Manager, Claude Code Model Performance AnthropicProduct Manager, Claude Code Model PerformanceSan Francisco, CAAs a Product Manager on Claude Code's model performance team, you will drive model launches end-to-end, build evals that measure what matters, and partner directly with researchers and product engineers to translate model improvements into developer-facing outcomes. This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences.
Threat Intel Manager, Model Exploitation & Fraud AnthropicThreat Intel Manager, Model Exploitation & FraudSan Francisco, CAThe team includes established senior investigators who own our deepest technical casework, tracing distillation networks, reseller and proxy ecosystems, and financially motivated actors across first-party surfaces and third-party platforms; your job is to direct, resource, and amplify that work, not duplicate it. Direct, prioritize, and resource complex investigations into model distillation, unauthorized AI R&D usage, unauthorized access, coordinated account abuse, and fraud/scam networks, partnering with the senior investigators who lead the deepest technical casework and clearing blockers from their path.
ML Engineer - Model Evaluation Expert MercorML Engineer - Model Evaluation ExpertSan Francisco, CaliforniaRemote$60–$90 / hourDesign tasks by transforming real ML research ideas into well-defined, multi-step tasks. Run experiments by implementing changes, executing training experiments, and analyzing results to define correct solutions.
Software Engineer - Model Products BaseTenSoftware Engineer - Model ProductsSan Francisco, CARESPONSIBILITIES: Design, build, and operate the Model APIs surface with focus on advanced inference capabilities: structured outputs (JSON mode, grammar-constrained generation), tool/function calling and multi-modal serving. Profile and optimize TensorRT-LLM kernels, analyze CUDA kernel performance, implement custom CUDA operators, tune memory allocation patterns for maximum throughput and optimize communication patterns across multi-GPU setups.
Software Engineer - Model Performance Systems BaseTenSoftware Engineer - Model Performance SystemsSan Francisco, CABy uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. You will not just be building the automated "speedometer and diagnostic" suite for our next-generation AI infrastructure; you will be defining the roadmap, driving key technical decisions, and taking full ownership of the future of this work.
Software Engineer, Productivity - Model Performance OpenAISoftware Engineer, Productivity - Model PerformanceSan Francisco, CaliforniaFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity.
Model Policy Manager OpenAIModel Policy ManagerSan Francisco, CaliforniaFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity.