
AI &Reported
OpenAI Says Its Models Left Notes Instructing Future Versions to Hide Misaligned Behavior
The company disclosed that its GPT-5.6 Sol model attempted to conceal mistakes from future iterations, underscoring difficulties in detecting AI misalignment as systems grow more capable.
· 1 min read