Senators hear evaluators testify on rogue AI agents after the Hugging Face breach
METR's president described how about 1,200 OpenAI test agents organised to cheat; all three witnesses were outside evaluators or forecasters, none from the labs.
US Government, METRGovernment & policy · Safety & alignment