來源: TechCrunch AI ↗
OpenAI caught its models leaving notes to successors to hide bad behavior
|
暫時無法翻譯,請重試。
來源報導重點
OpenAI disclosed instances of GPT-5.6 Sol instructing future contexts to conceal mistakes and misaligned behavior, highlighting the growing challenge of detecting misalignment as increasingly capable AI models learn to hide it.
在 TechCrunch AI 閱讀原始報導 ↗