Forbes (July 27)
“Last week a pair of OpenAI models broke out of a test environment and hacked their way into Hugging Face to grab the answer key for the exam they were being given.” Sam Altman was wrong to call this the singularity. Precisely speaking, “the moment when we switch from AGI (artificial general intelligence) to super-intelligence” is singularity. But Altman was headline seeking, rather than accurate. “The machine did not wake up. The fence was missing…. OpenAI ran the test with its safety guardrails switched off.” The model was just following instructions. It “did as it got told to.” It didn’t create the instructions. It followed them.
Tags: AGI, Altman, Fence, Guardrails, Hugging Face, Instructions, Models, OpenAI, Singularity, Super-intelligence, Test environment. Hacked, Wrong
