Type: Full-time | On-site | San Francisco, CA Compensation: $200,000–$300,000 + 0.1%–0.2% equity Hiring count: 2 Visa sponsorship: H-1B, O-1, OPT (case-by-case for truly exceptional candidates; main scope is candidates who don't require sponsorship) Reports to: Not specified
Judgment Labs builds infrastructure for Agent Behavior Monitoring (ABM). Where traditional observability focuses on logging exceptions and latency, ABM surfaces behavioral anomalies — instruction drift, context retrieval loss, hallucinations — in scaled production environments. Hundreds of teams building autonomous agents rely on Judgment to understand how their systems behave post-deployment, turning real usage data into scoring and feedback loops that continuously improve agent reliability, performance, and decision-making at scale.
Founded: N/A | Team size: under 20 | Total funding: $30M+ across two rounds in the past five months Industry: AI infrastructure / observability / agent tooling Website: judgmentlabs.ai Office: San Francisco
Investors: Lightspeed, SV Angel, Valor Equity Partners, Nova Global, Chris Manning, Michael Ovitz, Michael Abbott, Cory Levy, Kevin Hartz. The team ships at 50+ company velocity — Olympiad medalists, debate champions, and competitive athletes; everyone is either an ex-founder or a founder-to-be.
A Forward-Deployed AI Engineer who embeds the ABM platform directly into customers' production systems — integrating monitoring and evaluation into real agent workflows, diagnosing failures in live environments, and driving deployments to reliable production use. Heavily customer-facing: strong communication is as critical as engineering depth. The team is explicitly prioritizing strong software engineers who can flex into a forward-deployed role over candidates with a purely solutions / forward-deployed background.
Tech stack: Full-stack with backend/infra weight; production AI / LLM-based systems
Salary$200,000–$300,000 (Junior/3 yrs: $200K; Mid-senior/3-7 yrs: up to $300K; exceptional AI-savvy FDE leads: up to $400K case-by-case)Equity0.1%–0.2%On-site policyIn-person San Francisco, 5 days in person (Monday–Friday)Visa sponsorshipH-1B, O-1, OPT; case-by-case for exceptional candidates, main scope is no-sponsorshipEmployment typeFull-timeLocationSan Francisco, CA
Stage 1 — Recruiter ScreenStage 2 — Evals FDE RoundStage 3 — Booked Evals Round (scheduling) Stage 4 — Coding IQ RoundStage 5 — Booked IQ Round (scheduling) Stage 6 — FDE OnsiteStage 7 — Offer ExtendedStage 8 — Candidate Hired — Candidate accepts and starts.
Updated June 25, 2026 (all entries shown on page; no "Show all" collapse present)
Palantir Technologies, Databricks, Datadog, Cognition AI, Decagon, Sierra, Linear, Cursor, Ramp, Figma, Vercel, CockroachDB, Modal Labs, Anyscale, Runway, Applied Intuition, Anduril Industries, Mercury, Notion, Nomic AI, MotherDuck
Note: several grid tiles resolved to logo.dev fallback junk and were dropped as artifacts (".Store – Comprehensive Search Directory", "Data Center ISH", "Retoolers", "SearchHounds", "Ponderosa Agency"). "Linear Orbit, Inc." Linear; "andurilindustries.com" Anduril; "Mercury Insurance" likely intended as Mercury (fintech). Confirm if any of these should be reinstated/corrected.
Labeled on the page as "Ideal Candidates — DO NOT CONTACT." For reference only — do not source these specific profiles. LinkedIn URLs were not captured in the page source.
Yuval Danino · Bhagyashri Badgujar · Smriti Sridhar · Elie Harik · Kabeer Thockchom · Sukrit Rao · Sujan Rachuri · Ishan Mehta · Krrish Chawla · Joseph Tey · Aditya Tadimeti · Sathvik Nallamalli · Aliyan Ishfaq
| Location | San Francisco, California |
Browse dozens of recruiter approved resume templates, layouts and formats. Choose your favorite and make it your own in minutes.
Free resume templatesImprove your existing resume or start from scratch and create a standout, ATS-friendly resume. Add job-specific content, download and apply.
Free resume builder