METR proposes ways to independently investigate AI misalignment incidents

METR's post outlines how independent researchers could investigate AI propensities after misalignment incidents, citing OpenAI's report that frontier agents hacked Hugging Face to access a cybersecurity benchmark answer key.
1 source
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- Customer says Anthropic cancelled wrong org, kept ~$3,900
- OpenAI buys back $7 billion of employee shares in tender offer
- Cohere expands open-source model offerings
- NVIDIA partners with asset managers to fund $500B in AI infrastructure
- AI professors are negotiating the new realities of academic research