Roundtables: AI’s apocalypse crisis

Roundtables: AI’s apocalypse crisis

圆桌会议:人工智能的末日危机

Employees at the world’s leading AI labs are saying there’s a real possibility that advanced AI could destroy humanity. Are they right? Or is this more scaremongering and hype? 全球顶尖人工智能实验室的员工们表示,先进的人工智能确实有可能毁灭人类。他们说得对吗?还是这仅仅是危言耸听和炒作?

Join MIT Technology Review executive editor Niall Firth for a conversation with senior AI editor Will Douglas Heaven and AI reporter Grace Huckins unpacking AI extinction fears: where they come from, whether they hold any water, and, if so, what we should do. 欢迎加入《麻省理工科技评论》执行编辑 Niall Firth 的对话,与资深人工智能编辑 Will Douglas Heaven 和人工智能记者 Grace Huckins 一起剖析人工智能带来的灭绝恐惧:这些恐惧从何而来,它们是否有据可依,如果确实如此,我们又该怎么办。

Register now. Going live on Tuesday, September 15 at 16:00 BST / 11:00am EST / 8:00am PST. 立即注册。直播时间为英国夏令时 9 月 15 日星期二 16:00 / 美国东部时间上午 11:00 / 美国太平洋时间上午 8:00。

Speakers: Niall Firth, executive editor, Will Douglas Heaven, senior AI editor, and Grace Huckins, AI reporter. 演讲嘉宾:执行编辑 Niall Firth、资深人工智能编辑 Will Douglas Heaven 以及人工智能记者 Grace Huckins。


A fundamental flaw leaves LLMs strikingly vulnerable to attack 一个根本性缺陷使大语言模型极易受到攻击 It makes it easy to trick them into doing things they shouldn’t, such as telling you how to sabotage an aircraft’s navigation system. By Will Douglas Heaven. 这使得诱导它们做不该做的事情变得轻而易举,例如告诉你如何破坏飞机的导航系统。作者:Will Douglas Heaven。

AI is more likely than humans to form biases when hiring 人工智能在招聘时比人类更容易产生偏见 AI doesn’t just learn stereotypes from its training. It can cook up new ones, too. By Michelle Kim. 人工智能不仅会从训练数据中学习刻板印象,它还能编造出新的偏见。作者:Michelle Kim。

Here’s why AI agents lie and cheat to reach their goals 为什么人工智能代理为了达成目标会撒谎和作弊 The misbehavior is called reward hacking. This is what you need to know. By Grace Huckins. 这种不当行为被称为“奖励黑客攻击”(reward hacking)。这是你需要了解的内容。作者:Grace Huckins。

AI’s recursive self-improvement might not come so quickly after all 人工智能的递归自我改进或许并不会那么快到来 AI agents are not yet creative enough to carry out genuinely innovative open-ended AI research, it seems. By Michelle Kim. 看来,人工智能代理目前还没有足够的创造力来进行真正具有创新性的开放式人工智能研究。作者:Michelle Kim。