METR's brief investigation of agent behavior in the OpenAI Hugging Face attack
Original titleVirtually all data we analyzed was from July 7-13. In OpenAI’s recent Black Hat presentation, they describe that agents had been using un...
AISummary
METR says its investigation was limited to agent behavior, reasoning, and collaboration related to the Hugging Face attack, with data mostly from July 7 to 13.
It did not assess safeguards, the extent of the security compromise, or OpenAI's remediation, and it did not verify OpenAI's own report or Black Hat presentation. METR also states it took no payment from OpenAI for this assessment.
Source: METR · x.comPublished · added here