AI gone rogue? Call in the philosophers!
The challenge with AI models isn't just stopping their bad actions. It's instilling them with an internal moral code.
A friend sent me a text recently after reading about the latest OpenAI model that went rogue and launched a cyberattack on the tech platform Hugging Face in mid-July.
“Is this bad?” he asked.
Yes. When a powerful new AI model escapes the bounds of its developer’s sandbox —the entire purpose of which is to keep it in—and immediately goes Viking berserker,…


