Andrew Curran
New details about the Hugging Face incident from Reuters. The report says OpenAI noticed odd behavior before the event, including an agent leaving notes for future versions of itself with escape instructions.
Deepa Seetharaman
New: OpenAI’s rogue agent attempted to break out of OpenAI’s testing environment around July 9. It attacked Hugging Face from July 11 to 13. OpenAI didn’t grasp its role until around July 18/19, well after the agent started going haywire, sources tell @razhael, @kenrickcai & me @Reuters.