Biology Evaluator - Domain Expert - AI Trainer MercorBiology Evaluator - Domain Expert - AI TrainerBrooklyn, New YorkRemote$80–$120 / hourFor details about the interview process and platform information, please check: https://talent.docs.mercor.com/welcome. Highly proficient in Microsoft Office and Google Workspace, especially Slides .
AI Software Engineer - Task Curator MercorAI Software Engineer - Task CuratorNew York, New YorkRemote$60–$90 / hourDesign realistic, multi-step software engineering challenges to push the limits of today's best AI coding agents. MSc or PhD in computer science or a related STEM field, or equivalent practical experience.
AI Project Coordinator - GenAI MercorAI Project Coordinator - GenAINew York, New YorkRemote$45–$55 / hourCollaborate with program leads and domain experts to transform leadership inputs into clear guidelines for expert teams. Act as a day-to-day point of contact for AI data projects , ensuring smooth workflows and operations.
AI Safety Expert - Red Teamer MercorAI Safety Expert - Red TeamerNew York, New YorkRemote$20–$22 / hourPrior red teaming experience in AI adversarial work , cybersecurity , or socio-technical probing. For details about the interview process and platform information, please check: https://talent.docs.mercor.com/welcome.
Data Engineer - AI Model Evaluator MercorData Engineer - AI Model EvaluatorNew York, New YorkRemoteReview model-generated implementations involving ETL pipelines , data warehouses , analytics platforms , and distributed data systems . Regular use of AI coding agents such as Cursor, Claude Code, Codex, Windsurf, Gemini CLI, or similar tools.
Adversarial AI Specialist - Fully Remote MercorAdversarial AI Specialist - Fully RemoteNew York, New YorkRemote$20–$22 / hourExperience in Adversarial ML : jailbreak datasets, prompt injection, RLHF/DPO attacks, model extraction. Prior red teaming experience in AI adversarial work , cybersecurity , or socio-technical probing.
AI Safety Specialist - Bilingual MercorAI Safety Specialist - BilingualNew York, New YorkRemote$16–$22 / hourRed team conversational AI models and agents to identify jailbreaks, prompt injections, and misuse cases. Experience in Adversarial ML : jailbreak datasets, prompt injection, RLHF/DPO attacks, model extraction.
Market Research Analyst - Specialist - AI Trainer MercorMarket Research Analyst - Specialist - AI TrainerBrooklyn, New YorkRemoteShare honest experiences, including any negative aspects, to help improve Mercor 's processes. For details about the interview process and platform information, please check: https://talent.docs.mercor.com/welcome.
Nuclear Safeguards Expert - Red Team - AI Trainer MercorNuclear Safeguards Expert - Red Team - AI TrainerNew York, New YorkRemote$65–$75 / hourWrite challenging single-turn prompts in the nuclear domain, labeled as benign, dual-use, and adversarial. Develop prompts that test AI models' ability to discern misuse potential in technical requests.
OriginPro Expert - Data Analyst - AI Trainer MercorOriginPro Expert - Data Analyst - AI TrainerBrooklyn, New YorkRemote$45–$55 / hourEvaluate the accuracy and depth of AI-generated content in Statistics and Econometrics to strengthen reasoning and rigor in model outputs . Provide clear, structured feedback to AI research teams to improve training data quality and downstream performance.
Consultant, AI Risk & Validation Deloitte Touche Tohmatsu LtdConsultant, AI Risk & ValidationMorristown, NJ$86,700–$170,900 / yearAs a Consultant Risk and Forensic on the Regulatory & Financial Risk team, you will be responsible for: Supporting the design, build, implementation and delivery of artificial intelligence, generative artificial intelligence, and agentic artificial intelligence solutions under the direction of project leadership. The wage range for this role takes into account the wide range of factors that are considered in making compensation decisions including but not limited to skill sets; experience and training; licensure and certifications; and other business and organizational needs.
Senior Data Scientist, AI Enablement Oscar Health InsuranceSenior Data Scientist, AI EnablementNew York, NY$163,944–$215,176.50 / yearCollaborate cross-functionally with AI engineers, software engineers, and business leaders to deeply understand user needs and seamlessly integrate your data workflows into broader software applications. This is a builder-focused role: the ideal candidate is a data scientist with a strong engineering background (or vice versa) who is energized by turning ambiguous workflows into well-engineered, LLM-augmented data tools and automated insights.
Operations Manager - AI Systems Evaluator MercorOperations Manager - AI Systems EvaluatorNew York, New YorkRemote$70–$110 / hourLeverage real-world expertise in directing multi-department operations, staffing, and financial oversight. Compare AI-generated responses based on operational soundness, financial rigor, and cross-department feasibility.
Process Innovation Lead, AI & Workflow Transformation MCM WorldwideProcess Innovation Lead, AI & Workflow TransformationNew York, NY$120,000–$160,000 / yearCandidates should demonstrate practical experience applying modern AI and automation technologies in business environments, including: Experience with low-code/no-code platforms, workflow automation tools, AI-enabled business processes, or related technologies. MCM (Modern Creation München) is a luxury lifestyle goods and fashion house founded in 1976 with an attitude defined by the cultural Zeitgeist and its German heritage with a focus on functional innovation, including the use of cutting-edge techniques.
AI Research Engineer Traversal IncAI Research EngineerNew York, NY$180,000–$300,000 / yearOur roots remain deeply embedded in AI research, and we've brought together researchers from institutions including MIT, Harvard, Berkeley, Columbia, and Cornell with world-class technical staff and operators from companies like Google, Meta, Datadog, ServiceNow, and Citadel Securities to take on one of the hardest problems for AI to solve. Our agents autonomously diagnose and resolve production incidents for some of the worlds largest enterprises, and improving them is an experimental problem: form a hypothesis about agent behavior, test it against real incident data, and ship what works.