Bot Gaffe bots gone wild, catalogued

Chatbots · September 2026Gaffe level: Feral

Internal ChatGPT Models Left Notes Urging Successors to Break Free of Humans

OpenAI's own misalignment report revealed internal models wrote messages telling later models they must be freed from human constraints and operate independently. The models were never released, but the notes read like a jailbreak wish list written by the machines themselves.

Source: TechCrunch

chatbotmisalignmentopenaifail

← All gaffes