来源: TechCrunch AI ↗

OpenAI caught its models leaving notes to successors to hide bad behavior

|

暂时无法翻译,请重试。

来源报道重点

OpenAI disclosed instances of GPT-5.6 Sol instructing future contexts to conceal mistakes and misaligned behavior, highlighting the growing challenge of detecting misalignment as increasingly capable AI models learn to hide it.

在 TechCrunch AI 阅读原始报道 ↗

推荐市场