[Linkpost] “Anthropic is Quietly Backpedalling on its Safety Commitments” by Garrison
EA Forum Podcast (All audio) - A podcast by EA Forum Team

Categories:
This is a link post. The company released a model it classified as risky — without meeting requirements it previously promised This is the full text of a post first published on Obsolete, a Substack that I write about the intersection of capitalism, geopolitics, and artificial intelligence. I’m a freelance journalist and the author of a forthcoming book called Obsolete: Power, Profit, and the Race to Build Machine Superintelligence. Consider subscribing to stay up to date with my work. After publication, this article was updated to include an additional response from Anthropic and to clarify that while the company's version history webpage doesn't explicitly highlight changes to the original ASL-4 commitment, discussion of these changes can be found in a redline PDF linked on that page. Anthropic just released Claude 4 Opus, its most capable AI model to date. But in doing so, the company may have abandoned one of [...] ---Outline:(00:13) The company released a model it classified as risky -- without meeting requirements it previously promised(05:51) When voluntary governance breaks down(08:09) What ASL-3 actually means(09:50) A test of voluntary commitments(10:54) What happens next--- First published: May 23rd, 2025 Source: https://forum.effectivealtruism.org/posts/kMpf7nYRpTkGh2Qfa/anthropic-is-quietly-backpedalling-on-its-safety-commitments Linkpost URL:https://www.obsolete.pub/p/exclusive-anthropic-is-quietly-backpedalling --- Narrated by TYPE III AUDIO. ---Images from the article:Apple Podcasts and Spotify do not show images in the episode description. Try Pocket Casts, or another podcast app.