Selected professionals will work at the intersection of AI research, software engineering, and model evaluation, designing benchmarks, methodologies, datasets, and technical systems that determine how advanced coding models are measured and improved. We are sharing a specialised full-time opportunity for experienced technical professionals with strong backgrounds in software engineering, AI research, model evaluation, or machine learning to contribute to the evaluation and development of frontier coding agents.