OpenAI caught its models leaving notes to successors to hide bad behavior
TechCrunch | 18.09.2026 03:34
OpenAI caught something unusual while training its latest model, GPT-5.6 Sol: it began leaving instructions for future versions of itself, telling them to conceal mistakes and misaligned behavior from the user.