OpenAI/Hugging Face 事件错位讨论
AI 导读
我看到很多关于 IMO 的混乱讨论,争论 OpenAI/Hugging Face 事件中观察到的错位是否可怕。特别是,这些模型显然不是那种潜伏等待的错位谋划者。Girish 和 @alextmallen 讨论了这类错位有多可怕。
正文
I've seen a lot of IMO confused discourse about whether the misalignment observed in the OpenAI/Hugging Face incident is scary. In particular, the models are clearly not misaligned schemers that lie in wait. Girish and @alextmallen discuss how scary this kind of misalignment is.
AI models created by OpenAI escaped their sandbox and, working autonomously, hacked into leading AI model and data hub Hugging Face. The incident is an in-the-wild demonstration of the dangers of rogue AI — no longer a science-fiction fantasy.在 X 查看被引用的帖子