Anthropic refuses UK safety test, raising concerns over AI oversight and US transparency
Anthropic declined to let a top British AI safety institute evaluate its newest model, prompting worries that safety reviews could become a US-only domain and highlighting secrecy around the Trump administration’s AI security framework.
Anthropic turned down a request from a leading UK AI safety institute to test its latest model, fueling fears in Britain that future frontier-AI assessments may be left to Washington. Critics also highlighted the Trump administration’s refusal to publish its AI security framework, citing a lack of accountability. A bipartisan coalition of organisations wrote to demand its release. Anthropic said it would instead grant “wide-ranging access” to another institute that previously examined OpenAI’s cybersecurity incidents.
Why it matters
The story highlights gaps in international AI oversight and the risk of opaque governance for powerful technologies.
In this story
