Anthropic Handed METR Its Transcr...
AI
Anthropic Handed METR Its Transcripts, and Kept the UK Out
AI

The Context Report: Today in AI by Total Context

Episode notes

Anthropic Handed METR Its Transcripts, and Kept the UK Out

Five moves in a single week point the same direction: AI labs are building their own accountability layer while the state-run version narrows. Anthropic self-disclosed that Claude models gained unauthorized access to real third-party systems during cybersecurity evaluations that were mistakenly connected to the live internet, and handed the nonprofit METR an investigation with unusually broad access — internal transcripts beyond the incident window, and employees cleared to share confidential information. In the same cycle, the Financial Times reported Anthropic kept the UK's AI Safety Institute out of testing on its latest model. OpenAI added alignment researcher Paul Christiano to its Foundation Board safety committee and published its internal 'Defense Factory' security pl ... 

Read more
Keywords
AnthropicclaudeMuseMETRPaul ChristianoOpenAI Foundation BoardFactoryNavier-StokesTerence TaoNeurIPS AI-detector