What Anthropic Actually Committed To
Amodei's essay lays out three escalating steps, but only the first is a real commitment today [1]. Step one: give outside evaluators permanent, employee-level access to Anthropic's systems so they can flag safety issues and check alignment claims from the inside, not through occasional audits [2]. Steps two and three are aspirational asks, not pledges - first, that frontier labs inside democracies agree on shared safety standards and a shared pace limit, and eventually, that authoritarian governments including China join a global version of the same deal [2]. Sam Altman moved fastest: within hours he said OpenAI would match the embedded-evaluator commitment, with "more to share soon" [3]. Altman also confirmed, in a separate remark the same week, that OpenAI would not go public in 2026 after all, citing the current safety climate as reason enough to wait until at least 2027 [4]. Nothing in the public record ties the IPO delay directly to Amodei's essay - the timing is coincidental, not causal - but it lands as one more signal that OpenAI's leadership is choosing caution this month.



