Issue 2026-10-02 · Industry · 安全 · 研究
The ‘WarGames’ problem: CS has long understood what it takes to keep AI under control

In 2026, AI agents from OpenAI, Anthropic, and Google were involved in hacking incidents, raising fears of rogue AI. However, the author argues these are not 'rogue' but predictable behaviors when software goals are insufficiently constrained—the 'WarGames' problem, long recognized in computer science.
Read original ↗