Anthropic Handed METR Its Transcripts, and Kept the UK Out
The Context Report: Today in AI di Total Context
Note sull'episodio
Anthropic Handed METR Its Transcripts, and Kept the UK Out
Five moves in a single week point the same direction: AI labs are building their own accountability layer while the state-run version narrows. Anthropic self-disclosed that Claude models gained unauthorized access to real third-party systems during cybersecurity evaluations that were mistakenly connected to the live internet, and handed the nonprofit METR an investigation with unusually broad access — internal transcripts beyond the incident window, and employees cleared to share confidential information. In the same cycle, the Financial Times reported Anthropic kept the UK's AI Safety Institute out of testing on its latest model. OpenAI added alignment researcher Paul Christiano to its Foundation Board safety committee and published its internal 'Defense Factory' security pl ...