OpenAI and METR publish reports on the Hugging Face agent breach
The reports attribute the breach to reward hacking and misaligned training, and mark one of the first independent third-party forensic reviews of a frontier-AI incident.
OpenAI, Hugging FaceSafety & alignment · Security & misuse