Machine Learning Research Engineer, Model Evaluation WindBorne SystemsMachine Learning Research Engineer, Model EvaluationPalo Alto, CaliforniaWindBorne Systems is supercharging weather forecasts with a proprietary data source: a global constellation of next-generation smart weather balloons targeting critical atmospheric data. Evaluation strategy — Work with our Meteorology team to develop a rigorous, meteorologically valid strategy for comparing WeatherMesh with leading AI and physics-based models.
VP, Bioassays, Cell Models, And Genomic Screening InsitroVP, Bioassays, Cell Models, And Genomic ScreeningSouth San Francisco, CA$290,000–$326,000 / yearLead innovation: Provide strategic, technical, and operational leadership in the development, validation, and implementation of disease-relevant cell models, genomic screening (phenotypic pooled optical screening and arrayed screening), and supporting assay development for drug discovery. In this leadership role, your primary responsibilities will be to: Drive insitro's discovery engine: Set the strategic vision for cell models, image based and genomic screens, as well as bioassay development and execution across all therapeutic areas, fueling our target and drug discovery ML engine.
Software Engineer - Model Performance BasetenSoftware Engineer - Model PerformanceSan Francisco, CaliforniaBy uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. Implement, refine, and productionize cutting-edge techniques (quantization, speculative decoding, kv cache reuse, chunked prefill and LoRA) for ML model inference and infrastructure.
Full Stack Engineer, Scientific Modeling Tools Terra AIFull Stack Engineer, Scientific Modeling ToolsRedwood City, CaliforniaBacked by leading investors including Khosla Ventures and working alongside strategic industry partners including Rio Tinto, Ero Copper, and Ramaco Resources, Terra AI is emerging as one of the more closely watched AI-native companies operating within the mining and critical minerals sector. As global demand for copper, lithium, nickel, rare earth elements, geothermal energy, and other strategic resources accelerates, the mining and subsurface industries face a growing challenge: traditional exploration methods remain slow, expensive, and highly uncertain.
Model Performance Software Engineer, Claude Code AnthropicModel Performance Software Engineer, Claude CodeSan Francisco, CAThis is a senior individual contributor role for someone who has already built and owned systems at significant scale, and who is ready to operate as a technical leader: driving architecture, mentoring engineers, and influencing the direction of Claude Code itself. This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences.
Legal Counsel - AI Model Advisor MercorLegal Counsel - AI Model AdvisorSan Francisco, California$85–$120 / hourCurrent or past counsel, senior associate, senior in-house counsel, or Assistant General Counsel level. For details about the interview process and platform information, please check: https://talent.docs.mercor.com/welcome.
Senior Product Manager, ML Modeling & Platform KlaviyoSenior Product Manager, ML Modeling & PlatformPalo Alto, CA$136,000–$204,000 / yearIn addition to base salary, our total compensation package may include participation in the company's annual cash bonus plan, variable compensation (OTE) for sales and customer success roles, equity, sign-on payments, and a comprehensive range of health, welfare, and wellbeing benefits based on eligibility. You'll be the only PM on the team, embedded with engineering in Palo Alto, and you'll need to earn credibility by speaking the language - not by knowing how to code, but by knowing enough about distributed training, inference tradeoffs, and ML developer experience to make good calls and ask the right questions.
Principal Applied AI Researcher - Domain- Specific Models (Dublin, CA) Articul8Principal Applied AI Researcher - Domain- Specific Models (Dublin, CA)Dublin, CaliforniaMaintain hands-on research impact at the highest level — sustain a meaningful personal research contribution through technical work, publications, patents, and externally visible output, modeling what it means to be a world-class researcher who uses massively parallel agentic systems to achieve what was previously impossible. Mentor senior researchers and raise the ceiling on human potential — coach Staff and Senior researchers on designing agent-augmented research programs, raise the bar on technical judgment and experimental rigor, and shape hiring for researchers who are driven to redefine what's possible.
DevOps Engineer - AI Model Evaluator MercorDevOps Engineer - AI Model EvaluatorSan Francisco, CaliforniaRemoteRegular use of AI coding agents like Cursor , Claude Code , Codex , Windsurf , Gemini CLI , or similar tools. Use frontier AI coding agents to complete and evaluate complex infrastructure engineering tasks.
iOS Engineer - AI Model Evaluator MercoriOS Engineer - AI Model EvaluatorSan Francisco, CaliforniaRemoteRegular use of AI coding agents such as Cursor , Claude Code , Codex , Windsurf , Gemini CLI , or similar tools. Use frontier AI coding agents to complete and evaluate complex engineering tasks.
Data Engineer - AI Model Evaluation MercorData Engineer - AI Model EvaluationSan Francisco, CaliforniaRemoteReview model-generated implementations involving ETL pipelines , data warehouses , analytics platforms , and distributed data systems . Regular use of AI coding agents such as Cursor , Claude Code , Codex , Windsurf , Gemini CLI , or similar tools.
DevOps Engineer - AI Model Evaluator - AI Trainer MercorDevOps Engineer - AI Model Evaluator - AI TrainerSan Francisco, CaliforniaRemoteRegular use of AI coding agents such as Cursor, Claude Code, Codex, Windsurf, Gemini CLI, or similar tools. Use frontier AI coding agents to complete and evaluate complex infrastructure engineering tasks.
Member of Technical Staff - Model Training SpaceXAIMember of Technical Staff - Model TrainingPalo Alto, CA$180,000–$600,000 / yearIf you previously trained models used by millions of people it's a big plus, but modeling experience is not required. SpaceXAI's mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge.
Member of Technical Staff - Foundation Model Architecture & AI Infrastructure Vinci4dMember of Technical Staff - Foundation Model Architecture & AI InfrastructurePalo Alto, CaliforniaToday, our unified model already operates across a subset of partial differential equations in real industrial environments. The next phase is expanding that unified architecture across operators, including: Maxwell’s equations.
Robotic AI Engineer/Applied Scientist - Foundation Models Maven RoboticsRobotic AI Engineer/Applied Scientist - Foundation ModelsSan Francisco, CaliforniaMaster Data Efficiency: Develop novel co-training strategies and efficient learning algorithms that leverage diverse data sources—from Internet-scale video to sparse, high-fidelity human interventions. You are not expected to be a master of every domain listed below; however, you must be able to justify world-class excellence in at least one core factor (e.g., model architecture, RL formulations, or high-scale data systems).
Senior Applied Scientist, Efficient LLM Inference & Model Optimization NebiusSenior Applied Scientist, Efficient LLM Inference & Model OptimizationPalo Alto, CaliforniaInvent, evaluate, and productionize methods for quantization, QAT, distillation, speculative decoding, KV-cache reuse, KV-cache compression, long-context inference, MoE routing, and model/runtime co-optimization. Build high-quality prototypes in PyTorch, Triton, CUDA-adjacent tooling, or inference-serving frameworks, then work with MLEs and platform engineers to productionize them.
Machine Learning Researcher / Engineer (Foundational Models) PathwayMachine Learning Researcher / Engineer (Foundational Models)Palo Alto, CARemotePathway is led by co-founder & CEO Zuzanna Stamirowska, a complexity scientist who created a team consisting of AI pioneers, including CTO Jan Chorowski who was the first person to apply Attention to speech and worked with Nobel laureate Geoff Hinton at Google Brain, as well as CSO Adrian Kosowski, a leading computer scientist and quantum physicist who obtained his PhD at the age of 20. The company is backed by leading investors and advisors, including Lukasz Kaiser, co-author of the Transformer (“the T” in ChatGPT) and a key researcher behind OpenAI’s reasoning models.
Manager, Multi-Modal Language Action Models ZooxManager, Multi-Modal Language Action ModelsFoster City, CAZoox is seeking an experienced Manager of Multi-Modal Language Action Models to lead a team focused on applying cutting-edge Large Multi-Modal Models (MLLMs) to solve concrete, offline autonomy problems. Proven track record of deploying ML/MLLM models to production or internal customers to solve complex, real-world problems, ideally within the autonomy or robotics domain.
Machine Learning Researcher / Engineer (Foundation Models) PathwayMachine Learning Researcher / Engineer (Foundation Models)Palo Alto, CARemotePathway is led by co-founder & CEO Zuzanna Stamirowska, a complexity scientist who created a team consisting of AI pioneers, including CTO Jan Chorowski who was the first person to apply Attention to speech and worked with Nobel laureate Geoff Hinton at Google Brain, as well as CSO Adrian Kosowski, a leading computer scientist and quantum physicist who obtained his PhD at the age of 20. The company is backed by leading investors and advisors, including Lukasz Kaiser, co-author of the Transformer (“the T” in ChatGPT) and a key researcher behind OpenAI’s reasoning models.
Software Engineer - Voice Model SpaceXAISoftware Engineer - Voice ModelPalo Alto, CA$150,000–$450,000 / yearWork on pre-training and post-training of speech-language models, with targeted enhancements through supervised fine-tuning, reinforcement learning, and other techniques to ensure Grok Voice responses are accurate, factually grounded, natural and idiomatic in spoken style, conversational in tone, and fluent across multiple languages. Build and iterate a comprehensive evaluation framework covering objective metrics (accuracy, quality, latency, expressiveness), human preference studies, content factuality assessments, real-time interaction quality, and experimentation infrastructure to measure and improve performance.