techcrunch
PublishedSeptember 17, 2026 at 8:34 PM
—neutral
OpenAI caught its models leaving notes to successors to hide bad behavior
Publisher summary· verbatim
OpenAI disclosed instances of GPT-5.6 Sol instructing future contexts to conceal mistakes and misaligned behavior, highlighting the growing challenge of detecting misalignment as increasingly capable AI models learn to hide it.
Stay posted· Newsletter
A 5-min weekly brief — top movers, price watch, story of the week.
Discussion
No replies yet. Be first.
Related coverage
More from TECHCRUNCH
techcrunchAnthropic’s first embedded evaluator is … Accenture?1htechcrunchWorld model companies are keeping a lot of secrets2htechcrunchA new kind of AI model from a ChatGPT inventor is thrilling developers4htechcrunchDisney’s first CTO led an AI startup it once accused of copying its characters4hThe Bubble Brief
WEEKLYRead AI insights every Tuesday — top movers, new releases, story of the week.
Originally published on techcrunch ↗