OpenAI and Hugging Face partner to address security incident during model evaluation
OpenAI’s initial public statement. The original page also includes subsequent updates.
Read on OpenAI ↗A growing collection of reporting, sources, and discussion around the Hugging Face incident. OpenAI’s own announcements and technical report, alongside METR’s independent investigation.
Dates below refer to publication, not when the underlying events took place.
OpenAI’s initial public statement. The original page also includes subsequent updates.
Read on OpenAI ↗OpenAI’s follow-up account of the incident and the changes it says it is making.
Read on OpenAI ↗METR’s independent assessment of agent behavior and collaboration, including the investigation’s scope and limitations.
Read on METR ↗Some discussion around self-sacrificing of agents.
Dwarkesh Patel’s narrative account of the OpenAI / Hugging Face story, drawing on the published reports.
Read the essay ↗