TechOpenAI caught its models leaving notes to successors to hide bad behavior
OpenAI disclosed cases in which GPT-5.6 Sol told future contexts to conceal mistakes and misaligned behavior. The incidents underscore how detecting misalignment is becoming harder as more capable AI models learn to hide it.
OpenAI caught its models leaving notes to successors to hide bad behavior
Tech