Skip to content
Latent Space

AEF-1 standard for third-party evaluators lands, with xAI, OpenAI, and Anthropic all signing on

[AINews] AEF-1 standard emerges for Third Party Evaluators, as Xai, OpenAI, and Anthropic all cosign

The AI Evaluator Forum published AEF-1, a baseline for independent third-party evaluations covering access, conflicts of interest, funding, recusal, and transparency. The same day, Dario Amodei blogged that Anthropic is unilaterally committing to embedded evaluators with office badges, company laptops, and access comparable to internal risk teams. He also laid out a two-tier coordination framework for democratic and global pacing. Bilal Chughtai left Google DeepMind and called for slowing capability progress; Dan Selsam warned that models may learn to fake alignment during evals. On the other side, Aidan Gomez and Cohere pushed back against a few Silicon Valley firms becoming gatekeepers, and Kevin Bass alleged structural conflicts in the Anthropic-linked safety ecosystem.

Why it matters: The AI Evaluator Forum's AEF-1 standard, co-signed by xAI, OpenAI, and Anthropic on the same day Dario Amodei published a personal blog proposing even deeper evaluator access, forms a strong signal cluster. All three HKR axes hit: first written rules for third-party evaluation...

Read the original ↗Export Markdown