If Anthropic and OpenAI only wanted a moat, they would demand rules for everyone else and keep the doors shut. Instead they volunteered badges, desks, and employee-level access for outside evaluators. That is expensive, leaky, and hard to fake.
The real tell is not the essay. It is whether those evaluators ever publish something the labs did not want published. Until then, “sinister cartel” maybe just a vibe.
We Must Pace the Frontier: I’ve written a new essay on why the AI industry should slow down, with a three-part plan for doing so.
Anthropic is unilaterally committing to the first of these steps. We’ll provide third-party evaluators with permanent, employee-level access to our systems, so that they can verify adherence to our safety measures, report on incidents, and assess models’ alignment during training.
You can read the full post here:
darioamodei.com/post/we-must…