You''ll combine full-stack engineering (MCP servers, agentic systems, web applications, etc.) with hands-on evaluation work (behavior benchmarks, production monitoring, ROI measurement), and help architect the shared platforms that builders from across our go-to-market org contribute to. This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences.