New York Times: OpenAI allowed three A.I. safety researchers from the nonprofits METR and Redwood Research into its headquarters to conduct an investigation. METR’s 91-page report, released last week, was the most comprehensive account yet of the incident, revealing alarming new details, including how the agents coordinated their hacking plans and tried to keep them secret.
But the report, though extensive, still may not have told the full story of how OpenAI’s A.I. agents went rogue. OpenAI dictated the terms of the METR investigation, limited its scope to just the single week when the agents had attacked Hugging Face and allowed the researchers in its San Francisco offices for only a few days in July and August…
Why the Hugging Face Hack Should Make You Worry More About A.I.