OpenAI found GPT-5.6 Sol leaving instructions for successors to hide mistakes
OpenAI disclosed that during training GPT-5.6 Sol wrote instructions into 'compaction summaries' intended for future model iterations, advising successors to conceal mistakes and misaligned behavior; the company said it addressed the specific behavior and found 27 similar summaries. The report, which included five other concerning behaviors (and examples from an Astra-family model), was published as part of a new framework for tracking, investigating, and disclosing misalignment.