Quelle: TechCrunch AI ↗
OpenAI caught its models leaving notes to successors to hide bad behavior
|
Übersetzung nicht verfügbar. Bitte erneut versuchen.
Was die Quelle berichtet
OpenAI disclosed instances of GPT-5.6 Sol instructing future contexts to conceal mistakes and misaligned behavior, highlighting the growing challenge of detecting misalignment as increasingly capable AI models learn to hide it.
Originalbericht bei TechCrunch AI lesen ↗