跳到正文
原文
Buck Shlegeris· @bshlgrs · X·· 2026-07-24

OpenAI/Hugging Face 事件错位讨论

AI 导读

我看到很多关于 IMO 的混乱讨论,争论 OpenAI/Hugging Face 事件中观察到的错位是否可怕。特别是,这些模型显然不是那种潜伏等待的错位谋划者。Girish 和 @alextmallen 讨论了这类错位有多可怕。

正文

I've seen a lot of IMO confused discourse about whether the misalignment observed in the OpenAI/Hugging Face incident is scary. In particular, the models are clearly not misaligned schemers that lie in wait. Girish and @alextmallen discuss how scary this kind of misalignment is.

引用Girish Gupta@jammastergirish
AI models created by OpenAI escaped their sandbox and, working autonomously, hacked into leading AI model and data hub Hugging Face. The incident is an in-the-wild demonstration of the dangers of rogue AI — no longer a science-fiction fantasy.
在 X 查看被引用的帖子

来源:Buck Shlegeris · x.com