AEF-1 Baseline and Embedded Evaluators
The AI Evaluator Forum published AEF-1 to set baseline expectations covering access, conflicts of interest, funding relationships, recusal, and transparency. Alongside this framework, Anthropic committed to hosting embedded third-party evaluators such as METR, promising internal-style access including office desks, access badges, and company laptops to verify safety adherence and examine training pipelines. [1]
Industry Split Over Pacing and Governance
The release of the standard comes amid sharp disagreement across the industry over pacing progress versus control-first safety. Former Google DeepMind researcher Bilal Chughtai warned that progress may be outrunning alignment and advocated for pacing and transparency, while Cohere's Aidan Gomez argued against allowing a small group of Silicon Valley companies to become AI gatekeepers for governments. [1]
Sources
- 01