DevOps Engineer - AI Model Evaluator - AI Trainer MercorDevOps Engineer - AI Model Evaluator - AI TrainerSan Francisco, CaliforniaRemoteRegular use of AI coding agents such as Cursor, Claude Code, Codex, Windsurf, Gemini CLI, or similar tools. Use frontier AI coding agents to complete and evaluate complex infrastructure engineering tasks.
iOS Engineer - AI Model Evaluator MercoriOS Engineer - AI Model EvaluatorSan Francisco, CaliforniaRemoteRegular use of AI coding agents such as Cursor , Claude Code , Codex , Windsurf , Gemini CLI , or similar tools. Use frontier AI coding agents to complete and evaluate complex engineering tasks.
Data Engineer - AI Model Evaluation MercorData Engineer - AI Model EvaluationSan Francisco, CaliforniaRemoteReview model-generated implementations involving ETL pipelines , data warehouses , analytics platforms , and distributed data systems . Regular use of AI coding agents such as Cursor , Claude Code , Codex , Windsurf , Gemini CLI , or similar tools.
Model Maker Artech LLCModel MakerFoster City, CA$50–$62.92 / hourYou'll work on highly dynamic projects that come to the prototyping team, interfacing with both prototyping teammates and internal clients to support rapid iteration for vehicle development. You will both bring and grow a prototyping mindset in addition to various capabilities including (but not limited to) metalwork, composites, and additive manufacturing.
Model Designer OpenAIModel DesignerSan Francisco, CAFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity.
Research Program Manager - Adversarial Model Research OpenAIResearch Program Manager - Adversarial Model ResearchSan Francisco, CAFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. The Human Data team at OpenAI is responsible for identifying and mitigating risks in advanced AI systems by designing evaluations, surfacing vulnerabilities, and collaborating closely with researchers to strengthen model reliability and public trust.
Machine Learning Engineer, Model Evaluations (Speech LLM) - San Francisco PlaudMachine Learning Engineer, Model Evaluations (Speech LLM) - San FranciscoSan Francisco, CaliforniaWith a mission to amplify human intelligence, Plaud captures, structures, and compounds the intelligence generated in conversations — so humans can think better, decide faster, and execute with clarity. Can deeply partner with ML researchers to define exactly what "good" looks like for a Speech LLM, translating capabilities (like ASR robustness in noisy environments or TTS emotional steerability) into measurable benchmarks.
Solutions Architect - AI Model Specialist FriendliAISolutions Architect - AI Model SpecialistSan Francisco, CaliforniaOur infrastructure powers high-throughput, low-latency workloads for global organizations and integrates directly with Hugging Face, providing instant access to over 600,000 open-source models. You will work closely with our customers to integrate FriendliAI’s inference and agent frameworks into real-world products, enabling them to build and scale AI applications effectively.
Research Program Manager – Adversarial Model Research OpenAIResearch Program Manager – Adversarial Model ResearchSan Francisco, CaliforniaFor unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. The Human Data team at OpenAI is responsible for identifying and mitigating risks in advanced AI systems by designing evaluations, surfacing vulnerabilities, and collaborating closely with researchers to strengthen model reliability and public trust.
Senior Machine Learning Engineer - Foundation Model XPeng MotorsSenior Machine Learning Engineer - Foundation ModelSanta Clara, CA$174,720–$295,680 / yearYou will work closely with world-class researchers, perception and planning engineers, and infrastructure experts to design, train, and deploy large-scale multi-modal models that unify vision, language, and control. XPENG is a leading smart technology company at the forefront of innovation, integrating advanced AI and autonomous driving technologies into its vehicles, including electric vehicles (EVs), electric vertical take-off and landing (eVTOL) aircraft, and robotics.
PR Manager, Cloud And Model Partners NvidiaPR Manager, Cloud And Model PartnersSanta Clara, CAWhat you'll be doing: Support PR activities for NVIDIA's partnerships and new innovations with cloud service providers, neo cloud companies, model makers and other AI labs. Ways to stand out from the crowd: Experience navigating high-visibility partnerships and customer stories during news moments and in between.
Software Engineer- Model Performance Systems BaseTenSoftware Engineer- Model Performance SystemsSan Francisco, CABy uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. You will not just be building the automated "speedometer and diagnostic" suite for our next-generation AI infrastructure; you will be defining the roadmap, driving key technical decisions, and taking full ownership of the future of this work.
Product Manager, Claude Code Model Performance AnthropicProduct Manager, Claude Code Model PerformanceSan Francisco, CAAs a Product Manager on Claude Code's model performance team, you will drive model launches end-to-end, build evals that measure what matters, and partner directly with researchers and product engineers to translate model improvements into developer-facing outcomes. This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences.
Research Engineer, Production Model Post-Training AnthropicResearch Engineer, Production Model Post-TrainingSan Francisco, CAThis research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences. As a Research Engineer on our Post-Training team, you'll train our base models through the complete post-training stack to deliver the production Claude models that users interact with.
Research Product Manager, Model Behaviors AnthropicResearch Product Manager, Model BehaviorsSan Francisco, CAThis research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences. We offer competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space in which to collaborate with colleagues.
Staff Machine Learning Engineer - Foundation Model XPeng MotorsStaff Machine Learning Engineer - Foundation ModelSanta Clara, CA$215,280–$364,320 / yearDesign, train, and deploy large deep learning models that can leverage the vast amount of labeled and unlabeled data from a fleet of million vehicles. Within the range, individual pay is determined by work location and additional factors, including job-related skills, experience, and relevant education or training.
Threat Intel Manager, Model Exploitation & Fraud AnthropicThreat Intel Manager, Model Exploitation & FraudSan Francisco, CAThe team includes established senior investigators who own our deepest technical casework, tracing distillation networks, reseller and proxy ecosystems, and financially motivated actors across first-party surfaces and third-party platforms; your job is to direct, resource, and amplify that work, not duplicate it. Direct, prioritize, and resource complex investigations into model distillation, unauthorized AI R&D usage, unauthorized access, coordinated account abuse, and fraud/scam networks, partnering with the senior investigators who lead the deepest technical casework and clearing blockers from their path.
Senior Staff Machine Learning Engineer – Autonomous Driving Foundation Models XPeng MotorsSenior Staff Machine Learning Engineer – Autonomous Driving Foundation ModelsSanta Clara, CA$244,140–$413,160 / yearXPENG is a leading smart technology company at the forefront of innovation, integrating advanced AI and autonomous driving technologies into its vehicles, including electric vehicles (EVs), electric vertical take-off and landing (eVTOL) aircraft, and robotics. Key Responsibilities: Architectural Leadership: Lead the design of end-to-end VLA architectures, bridging multi-modal perception with high-level linguistic reasoning and precise action generation.
Software Engineer - Model Products BaseTenSoftware Engineer - Model ProductsSan Francisco, CARESPONSIBILITIES: Design, build, and operate the Model APIs surface with focus on advanced inference capabilities: structured outputs (JSON mode, grammar-constrained generation), tool/function calling and multi-modal serving. Profile and optimize TensorRT-LLM kernels, analyze CUDA kernel performance, implement custom CUDA operators, tune memory allocation patterns for maximum throughput and optimize communication patterns across multi-GPU setups.
ML Engineer - Model Evaluation Expert MercorML Engineer - Model Evaluation ExpertSan Francisco, CaliforniaRemote$60–$90 / hourDesign tasks by transforming real ML research ideas into well-defined, multi-step tasks. Run experiments by implementing changes, executing training experiments, and analyzing results to define correct solutions.