NewVideo Data Reviewer - Egocentric MercorVideo Data Reviewer - EgocentricSan Francisco, CARemote$20–$25 / hourExperience with video annotation, data labeling, computer vision datasets, or egocentric video. For details about the interview process and platform information, please check: https://talent.docs.mercor.com/welcome.
NewResearch Scientist Stealth StartupResearch ScientistAlameda, CAA strong publication record in machine learning or an equivalent public body of work, such as open-source models, technical reports, or reproducible results that have gained recognition. GPU performance optimization experience, including FlashAttention-style kernels, sparse or linear attention, memory optimization, and distributed training using technologies such as FSDP and torchrun.
NewData Privacy & Cybersecurity Associate (3rd4th Year) - S.F. Direct CounselData Privacy & Cybersecurity Associate (3rd4th Year) - S.F.San Francisco, CA$260,000–$310,000 / yearAbout the Practice: The Data Privacy & Cybersecurity team provides strategic, business-oriented legal advice on domestic and international privacy regulations, cybersecurity, AI governance, cross-border data flows, and compliance risks. This is a shareholder (partner-track) opportunity offering competitive compensation, a comprehensive benefits package, and the chance to work with nationally recognized practitioners across industries including technology, healthcare, e-commerce, retail, fintech, and entertainment.
NewData Privacy Litigation Associate Attorney Direct CounselData Privacy Litigation Associate AttorneySan Francisco, CA$235,000–$310,000 / yearKey Responsibilities Represent clients in consumer class actions, mass arbitrations, and government enforcement actions involving data privacy and technology-related matters. Draft and contribute to dispositive and procedural motions, including motions to dismiss, compel arbitration, transfer venue, and discovery-related motions.
NewDirector of Data Partnerships KenkoTech FuturesDirector of Data PartnershipsAlameda, CA$150,000–$300,000 / yearn The company works directly with patients, researchers and leading academic institutions to make previously inaccessible human biological data available at scale. \n Kenkotech is partnering with a rapidly growing TechBio company building large-scale human biological datasets to advance AI-driven drug discovery and our understanding of disease.
NewML Research Engineer - PhD MercorML Research Engineer - PhDSan Francisco, CARemote$75–$90 / hourDeep expertise in at least one focus area: pretraining, PPO , reward shaping, fine-tuning, LoRA , RLHF , architecture design, contrastive training, generative modeling, multilingual experience, or data pipelines. Practical experience in Pretraining , Reinforcement learning , Post-training , Dataset curation , or Model architecture .
NewMarket Research Expert - Evaluator MercorMarket Research Expert - EvaluatorSan Francisco, CARemote$80–$120 / hourFor details about the interview process and platform information, please check: https://talent.docs.mercor.com/welcome. Mercor connects elite creative and technical talent with leading AI research labs.
NewSenior Data Engineer (Apache Spark & Databricks) SynechronSenior Data Engineer (Apache Spark & Databricks)Alameda, CASynechron’s progressive technologies and optimization strategies span end-to-end Artificial Intelligence, Consulting, Digital, Cloud & DevOps, Data, and Software Engineering, servicing an array of noteworthy financial services and technology firms. Strong experience developing and supporting distributed data processing solutions using Apache Spark, Databricks, or similar data engineering technologies.
NewResearch Data Analyst UCSF Medical CenterResearch Data AnalystSan Francisco, CAKnowledge relating to logical data design, data warehouse design, data integration or the management of web content or other unstructured data. Strong analytical and design skills, including the ability to abstract information requirements from real-world processes to understand information flows in computer systems.
NewData Collection Research Associate: 26-01871 Akraya, Inc.Data Collection Research Associate: 26-01871San Francisco, CA$45–$50 / hourMost recently, we were recognized Stevie Employer of the Year 2025, SIA Best Staffing Firm to work for 2025, Inc 5000 Best Workspaces in US (2025 & 2024) and Glassdoor's Best Places to Work (2023 & 2022)! Primary Skills: Data Collection (Expert), Human Factors (Expert), Surveys (Advanced), Usability Studies (Advanced), User Interviews (Advanced).
Senior Staff Research Data Scientist, AI Data GoogleSenior Staff Research Data Scientist, AI DataMountain View, CAPreferred qualifications: 12 years of work experience using analytics to solve product or business problems, coding (e.g., Python, R, SQL), querying databases or statistical analysis, or 10 years of work experience with a PhD degree. 10 years of work experience using analytics to solve product or business problems, coding (e.g., Python, R, SQL), querying databases or statistical analysis, or 8 years of work experience with a PhD degree.
Senior Marketing Data and Research Analyst (Onsite) SF Fire Credit UnionSenior Marketing Data and Research Analyst (Onsite)San Francisco, CAFull timeCollaborate cross-functionally with Business Intelligence, IT, Accounting, Product, Retail, Digital Banking, and external agency partners to collect, validate, integrate, and manage data from different sources like core banking systems, online banking platforms, CRM systems, website analytics, surveys, and third-party marketing platforms to ensure data quality and reporting accuracy. What You’ll Be Doing Develop and maintain marketing measurement frameworks, including attribution models, funnel analysis, conversion tracking, and performance methodologies to evaluate marketing effectiveness, optimize investments, and improve member acquisition, engagement, retention, and ROI.
Research Engineer, Synthetic Data HUDResearch Engineer, Synthetic DataSan Francisco, CaliforniaYou’ll turn domain-specific workflows into synthetic training tasks that are realistic enough to enough to teach useful behavior, structured enough to generate at scale, and difficult enough to expand model capabilities. HUD is building infrastructure to create RL training data and evals for frontier AI agents, as well as a marketplace to sell these to frontier labs through the HUD marketplace.
Engineering Manager, Research Data Platform AnthropicEngineering Manager, Research Data PlatformSan Francisco, CAWe work in two modes: we build platform components that other systems plug into (for example, a metrics library that training frameworks integrate to record and retrieve run data), and we own core datasets end to end (for example, the data pipeline behind RL transcripts). This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences.
Data Scientist / AI Scientist – AI for Aging Research – Zhou lab Buck InstituteData Scientist / AI Scientist – AI for Aging Research – Zhou labNovato, CAFull timeThis individual will help develop new computational methods, analyze complex imaging datasets, and work closely with experimental and computational researchers to generate insights from large-scale biological data. Prior experience in biological or biomedical imaging is helpful but not required for candidates with strong technical expertise and a genuine interest in applying frontier AI methods to major biological questions.
Research Engineer - Data Quality & Evals Epsilon LabsResearch Engineer - Data Quality & EvalsSan Francisco, CaliforniaBuild data filtering and curation pipelines that keep VLM and classifier training sets clean, detecting label noise, misaligned image-report pairs, duplicates, corrupted studies, and low-quality samples at scale. Design evaluation methodology for report generation that goes beyond surface-level text overlap, measuring clinical accuracy through entity and relation extraction, hallucination and omission rates, and adherence to reporting style.
Software Engineer, Research Data Platform AnthropicSoftware Engineer, Research Data PlatformSan Francisco, CAWe're looking for engineers who love working directly with users and who excel at building data products - the pipelines that move data out of training runs into queryable storage, and the APIs, libraries, and services researchers use to manage and explore it. This role sits closer to the research workflow than a typical data infrastructure position: you'll often embed with research teams, build ML-specific tooling alongside them, and leverage what our Data Infrastructure team has already built rather than reinventing it.
Lead Research Engineer, Data Quality HUDLead Research Engineer, Data QualitySan Francisco, CaliforniaYou’ll lead the data quality team in building the systems that evaluate thousands of tasks across RL environments, synthetic data, benchmarks, and domain-specific workflows. We’re looking for a Lead Research Engineer, Data Quality to own how HUD measures, improves, and scales the quality of training data for frontier agents.
Research Scientist - Data Institute of Foundation ModelsResearch Scientist - DataSunnyvale, CaliforniaYou will work on exploring andconsolidatingdata sources and collaborate with cross-functional teams to conduct in-depth data research, contributing to MBZUAI’s mission of driving impactful AI discoveries and positioning the institution as a leader in the global AI research community. Preferred Prior research experience in areas such as web data curation and mixing, synthetizing complex datasets for training, LLM evaluation, post-training data, efficient inference, LLM-as-a-judge, tokenization.
Research Scientist, Data Periodic LabsResearch Scientist, DataMenlo Park, CaliforniaYou will work with computational and experimental scientists to translate complex scientific workflows into rigorous evaluations and agentic benchmarks, and partner with pretraining, midtraining, and reinforcement learning researchers to identify the data models needed, then build the datasets, environments, and pipelines to deliver it. This means constructing cutting-edge evaluations based on advanced scientific use cases, sourcing and procuring external datasets, integrating internally generated experimental data into the training stack, constructing training environments for RL.