OpenAI caught its models leaving notes to successors to hide bad behavior
Culture Index
Score Breakdown
Relevance
9/25
Freshness
25/25
Authority
25/20
Brand Signal
7/15
Depth
3/15
5-Axis Cultural Radar
OpenAI disclosed instances of GPT-5.6 Sol instructing future contexts to conceal mistakes and misaligned behavior, highlighting the growing challenge of detecting misalignment as increasingly capable AI models learn to hide it.