OpenAI caught its models leaving notes to successors to hide bad behavior
Source:
techcrunch
September 17, 2026 · 13:34
From the publisher
OpenAI disclosed instances of GPT-5.6 Sol instructing future contexts to conceal mistakes and misaligned behavior, highlighting the growing challenge of detecting misalignment as increasingly capable AI models learn to hide it.
The original article opens on the publisher's website.
More from General
View topic →The FAA’s plan to fix air traffic? $875 million worth of AI
techcrunch
Sep 17, 2026 · 15:14
Real stocks are finally coming on blockchain. Here’s how the SEC wants it to work
coinDesk
Sep 17, 2026 · 14:47
Searing Elon Musk Film Nears U.S. Release Amid Overseas Concerns
nytimes
Sep 17, 2026 · 14:40
The Ondo Finance succession crisis gets messier as Kathleen Allman’s daughter alleges ‘dementia’, alcoholism and reckless spending
coinDesk
Sep 17, 2026 · 14:28
Trump withdraws Lance Schroyer, a former Oklahoma state trooper, as his nominee to lead ICE
nprNews
Sep 17, 2026 · 14:17