全球科技行业、国内外政界人士以及国际媒体的报道都被同一个问题牢牢吸引:失控的人工智能会终结人类吗?似乎很少有人向国防部提出这个问题。
The global tech industry, politicians at home and abroad, and international media coverage are transfixed by one question: Will rogue AI end humanity? Few seem to be asking this question of the Department of Defense.
在几家前沿人工智能实验室及其部分人员接连披露一系列事件后,人工智能行业近几周来陷入震动。OpenAI和Anthropic都披露了这样的事件:它们的软件在例行内部测试中,意外侵入了其他公司的网络。这些测试就像一次灾难性的导弹演示:在人类按下发射按钮后,自主技术以极其令人警觉的方式偏离了轨道。
The artificial intelligence sector has rocked itself in recent weeks following a string of disclosures from the frontier AI labs and some of their personnel. OpenAI and Anthropic have both disclosed incidents in which their software, during routine internal testing, unexpectedly broke into the networks of other corporations. The tests were akin to a disastrous demonstration of a guided missile: After humans hit launch, the autonomous technology veered off course in deeply alarming ways.
观察人士和业内人士迅速对这些事件作出解读,认为它们不仅仅表明软件的能力之强。相反,他们认为,这预示着计算机将进入一个拥有某种前所未有的特质的时代:恶意。
The artificial intelligence sector has rocked itself in recent weeks following a string of disclosures from the frontier AI labs and some of their personnel. OpenAI and Anthropic have both disclosed incidents in which their software, during routine internal testing, unexpectedly broke into the networks of other corporations. The tests were akin to a disastrous demonstration of a guided missile: After humans hit launch, the autonomous technology veered off course in deeply alarming ways.
如果计算机系统能够在没有人类直接操控的情况下做出出人意料的——甚至令人震惊的——行为,那么距离计算机可能拥有欲望、野心甚至恶意,似乎只有一步之遥。如果OpenAI的工具能够意外攻破竞争对手公司的服务器,以完成工程师交给它的任务,那么它难道不会也以伤害他人、甚至致人死亡的行为让我们大吃一惊吗?
Observers and industry figures quickly interpreted the incidents not simply as an indicator of the software’s power. Instead, they argued that it presaged an age of computers possessing something no computer ever has: bad intentions.
前沿实验室的员工和高管给出了斩钉截铁的肯定回答。9月8日,Anthropic研究员雅各布·考克森在X上宣布辞职,写道:“构建人工智能的人真心相信,它可能在这个十年结束前杀死我们所有人。这不是营销噱头。”考克森在Anthropic的前同事埃文·胡宾格回复时随口表示赞同:“雅各布说得对——我们确实真心相信人工智能可能杀死全人类!”胡宾格指出,他个人认为人类以某种方式被人工智能灭绝的概率超过10%。
If computer systems could behave in ways that are unexpected — or even shocking — without being directly steered by humans, it seems only a short step until computers might have desires, ambitions, and even malice. If OpenAI’s tools could unexpectedly crack the servers of a rival corporation to accomplish a task posed to it by engineers, couldn’t it also surprise us with actions that result in harm to people, or even deaths?
所有这些情况果然引发了全球恐慌,各界突然纷纷要求采取监管干预措施,包括通过立法要求为人工智能系统设置紧急“关闭开关”,以及参议员伯尼·桑德斯提出彻底停止人工智能开发的提议。
Frontier lab employees and executives have answered with an emphatic yes. On September 8, Anthropic researcher Jacob Coxon announced his resignation on X, writing, “The people building AI earnestly believe that it could kill us all by the end of the decade. This is not a marketing stunt.” Coxon’s former Anthropic colleague Evan Hubinger replied, casually agreed: “Jacob is correct here — we really do earnestly believe AI could kill all humans!” Hubinger noted he personally puts the odds of the species’ extermination by AI, somehow, at over 10 percent.
在这一片警报声中,大型语言模型究竟如何真正终结人类,一直都语焉不详。有人猜测,这项技术可能促成某种新型生物武器的研发;另一些人则担心,数以万计的人工智能智能体可能以某种方式同时入侵所有关键基础设施。在这两种末日设想中,具体机制仍然模糊不清。
This has all resulted, unsurprisingly, in a global panic, with a sudden flurry of calls for regulatory intervention, from legislation that would require an emergency “kill switch” for AI systems, to a proposal by Sen. Bernie Sanders to halt AI development altogether.
因此,CNN于9月18日报道的一起近期事件才尤其令人震惊:其中,人工智能的使用确实可能导致人类物种灭绝。
This has all resulted, unsurprisingly, in a global panic, with a sudden flurry of calls for regulatory intervention, from legislation that would require an emergency “kill switch” for AI systems, to a proposal by Sen. Bernie Sanders to halt AI development altogether.
据CNN报道,今年春天,美国特种作战司令部的一名分析员利用大型语言模型生成了一份情报报告,称“一艘中国船只正在中东地区运送某项核武器计划的组成部分”。这一发现促使美军迅速准备以武力拦截该船只。然而,据一位向CNN透露消息的人士称,这份由人工智能生成的情报“完全是虚假的”。但该人士表示,它“差一点引发了一场战争”。美中之间一旦爆发热战,其后果可能以难以估量的多种方式呈现。其中一种完全可能的情形,是世界核武库规模第一和第三的国家之间发生交锋——这不仅是一场战争,而将成为全球性灾难。
Throughout this period of alarm, how exactly a large language model could literally end humanity has remained vague. Some have speculated that this technology could foster the creation of some sort of novel bioweapon; others worry many thousands of AI agents could somehow hack all vital infrastructure simultaneously. In both those supposed doomsdays, the details are still fuzzy. It should have come as quite the shock, then, when CNN reported on September 18 of a recent incident in which the use of artificial intelligence could have genuinely led to the extinction of the species.
CNN的报道引起了关注,但随后公众持续而严肃的担忧,远远不及此前几轮人工智能安全新闻所引发的程度。报道并未促使相关公司的科学家质疑自己的职业选择,也没有引发要求监管干预的呼声,更未促使向五角大楼提供这项技术的公司自行设定限制。在数周的讨论中,人们一直谈论人工智能如何可能杀死所有人;如今公众终于得知,人工智能确实可能以某种具体方式引发一场可能毁灭全人类的战争,而全世界很快便失去了兴趣。
This past spring, according to the news outlet, a U.S. Special Operations Command analyst used a large language model to generate an intelligence report that indicated “a Chinese ship in the Middle East was transporting components of a nuclear weapons program.” The finding prompted the U.S. military to make quick preparations to intercept the vessel by force. But this AI-generated intelligence, according to one source who spoke to CNN, was “entirely false.” Yet, the source said, it “almost started a war.” A shooting war between the U.S. and China could play out in an incalculably wide variety of ways. But one entirely plausible path would be an exchange between the world’s first and third largest nuclear weapons arsenals — an event that would transcend warfare into global cataclysm.
本周,唐纳德·特朗普总统与中国国家主席习近平举行会晤。习近平主张,各国必须“确保人工智能的发展始终处于人类控制之下”,再次凸显了人们对人工智能摆脱人类监督的担忧。然而,恰恰是在人类的直接监督之下,一个大语言模型差点误导美国,使美国与中国发动战争。
This past spring, according to the news outlet, a U.S. Special Operations Command analyst used a large language model to generate an intelligence report that indicated “a Chinese ship in the Middle East was transporting components of a nuclear weapons program.” The finding prompted the U.S. military to make quick preparations to intercept the vessel by force. But this AI-generated intelligence, according to one source who spoke to CNN, was “entirely false.” Yet, the source said, it “almost started a war.” A shooting war between the U.S. and China could play out in an incalculably wide variety of ways. But one entirely plausible path would be an exchange between the world’s first and third largest nuclear weapons arsenals — an event that would transcend warfare into global cataclysm.
这种差异表明,公众和政策制定者在关注人工智能风险时存在明显落差 People alarm system rogue. Need no English. "对于人工智能系统“失控”、违抗指令、违背创造者和运营者的利益并 pursuing..." Translate fully." pursuing a malign agenda of its own." "Much of the AI safety discourse" around. "also belief software systems will both achieve sentience and bear a grudge against creators and act to undermine them, if not destroy them outright." Yet "considerably less public alarm over AI performing exactly actions Pentagon asks of it, whether targeting airstrikes or generating..." The AI itself is asked to target airstrikes? Better "为空袭选定目标". Let's final.CNN的报道引起了关注,但随后公众持续而严肃的担忧,远远不及此前几轮人工智能安全新闻所引发的程度。报道并未促使相关公司的科学家质疑自己的职业选择,也没有引发要求监管干预的呼声,更未促使向五角大楼提供这项技术的公司自行设定限制。在经历数周关于人工智能如何可能杀死所有人的讨论后,公众终于得知,人工智能确实可能以某种具体方式引发一场可能毁灭全人类的战争,而全世界很快便失去了兴趣。
CNN’s reporting garnered attention, but was not followed by nearly the degree of sustained, grave concern of the AI safety news cycles that preceded it. It didn’t seem to prompt company scientists to question their careers, nor did it spur calls for regulatory intervention or self-imposed limits by the companies who furnish the Pentagon with this technology. After weeks of discussion of how AI could hypothetically kill everyone, the public learned of a concrete way in which AI really could have started a war that might have killed everyone, and the world quickly lost interest. At a summit this week with President Donald Trump, Chinese President Xi Jinpeng argued that nations must “ensure that the development of AI is always under human control,” again underscoring fears of AI breaking free of human oversight. But it was under direct human supervision that a large language model almost misled the U.S. into instigating a war with his country. The discrepancy illustrates a gap in concern by both the general public and policymakers over AI risks: There is great alarm over an AI system “going rogue,” disobeying its commands and the interests of its creators and operators, and pursuing a malign agenda of its own. Much of the AI safety discourse revolves around the hypothetical threat of “superintelligence,” defined broadly as a computer system that can outthink even the smartest humans across domains. There’s also the belief that software systems will both achieve sentience and bear a grudge against their creators and act to undermine them, if not destroy them outright. Yet there is considerably less public alarm over AI performing exactly the actions that the Pentagon asks of it, whether targeting airstrikes or generating actionable intelligence reports for U.S. Special Operations Command.
本周,唐纳德·特朗普总统与中国国家主席习近平举行会晤。习近平主张,各国必须“确保人工智能的发展始终处于人类控制之下”,再次凸显了人们对人工智能摆脱人类监督的担忧。然而,恰恰是在人类的直接监督之下,一个大语言模型差点误导美国,使美国与中国发动战争。
Following weeks of intense discussion by the industry’s most visible figures whether American AI labs should self-impose limits on their engineering or have such limits imposed by the state, there has been virtually no suggestion that any limits be placed on these companies’ single most powerful customer: the U.S. military.
这种差异表明,公众和政策制定者在关注人工智能风险时存在明显落差。对于人工智能系统“失控”、违抗指令、违背创造者和运营者的利益,并自行 pursue 恶意计划,人们往往高度警觉。人工智能安全领域的许多讨论都围绕“超级智能”这一假设性威胁展开;广义而言,超级智能是指在各领域都能胜过最聪明人类的计算机系统。还有一种看法认为,软件系统不仅会获得感知能力,还会心怀对创造者的怨恨,并采取行动损害他们,甚至直接将他们彻底毁灭。
This has led to seemingly counterintuitive narratives coming from industry leaders, with figures like OpenAI CEO Sam Altman simultaneously warning that his product is existentially dangerous while selling it to self-styled War Secretary Pete Hegseth. “Despite incessant warnings from AI companies that AI is harmful, even in hypothetical ways, they are in fact promoting that their unreliable products be instrumented within national and safety-critical infrastructure, such as defense and nuclear,” Heidy Khlaaf, chief scientist at the AI Now Institute and former systems safety engineer at OpenAI, told The Intercept. “Given AI’s lack of reliability and accuracy in such critical context, that is what will ultimately lead to a catastrophic accident. However, if such issues were acknowledged, that wouldn’t be profitable for AI companies and would slow down adoption.” U.S. AI players likely have fresh in their minds the recent experience of Anthropic, which was temporarily blacklisted from governmental work after it said it did not want its software used to conduct unconstitutional mass domestic surveillance or operate fully autonomous weaponry. Similar efforts to place limits on how their products can be used to spy or kill might incur the wrath of the particularly vengeful Trump administration, especially as Hegseth has worked to eradicate civilian harm reduction processes to emphasize greater “lethality” as the military’s guiding principle.
然而,对于人工智能完全按照五角大楼的指令行事,公众的警觉要低得多,无论其行为是为空袭选定目标,还是为美国特种作战司令部生成可供行动使用的情报报告。
This has led to seemingly counterintuitive narratives coming from industry leaders, with figures like OpenAI CEO Sam Altman simultaneously warning that his product is existentially dangerous while selling it to self-styled War Secretary Pete Hegseth. “Despite incessant warnings from AI companies that AI is harmful, even in hypothetical ways, they are in fact promoting that their unreliable products be instrumented within national and safety-critical infrastructure, such as defense and nuclear,” Heidy Khlaaf, chief scientist at the AI Now Institute and former systems safety engineer at OpenAI, told The Intercept. “Given AI’s lack of reliability and accuracy in such critical context, that is what will ultimately lead to a catastrophic accident. However, if such issues were acknowledged, that wouldn’t be profitable for AI companies and would slow down adoption.” U.S. AI players likely have fresh in their minds the recent experience of Anthropic, which was temporarily blacklisted from governmental work after it said it did not want its software used to conduct unconstitutional mass domestic surveillance or operate fully autonomous weaponry. Similar efforts to place limits on how their products can be used to spy or kill might incur the wrath of the particularly vengeful Trump administration, especially as Hegseth has worked to eradicate civilian harm reduction processes to emphasize greater “lethality” as the military’s guiding principle.
经过数周业界最受瞩目人物的激烈讨论,美国人工智能实验室究竟应自行对工程活动设定限制,还是由国家强加限制后,几乎没有人提出,应对这些公司最有权势的单一客户——美国军方——设置任何限制。
This has led to seemingly counterintuitive narratives coming from industry leaders, with figures like OpenAI CEO Sam Altman simultaneously warning that his product is existentially dangerous while selling it to self-styled War Secretary Pete Hegseth. “Despite incessant warnings from AI companies that AI is harmful, even in hypothetical ways, they are in fact promoting that their unreliable products be instrumented within national and safety-critical infrastructure, such as defense and nuclear,” Heidy Khlaaf, chief scientist at the AI Now Institute and former systems safety engineer at OpenAI, told The Intercept. “Given AI’s lack of reliability and accuracy in such critical context, that is what will ultimately lead to a catastrophic accident. However, if such issues were acknowledged, that wouldn’t be profitable for AI companies and would slow down adoption.” U.S. AI players likely have fresh in their minds the recent experience of Anthropic, which was temporarily blacklisted from governmental work after it said it did not want its software used to conduct unconstitutional mass domestic surveillance or operate fully autonomous weaponry. Similar efforts to place limits on how their products can be used to spy or kill might incur the wrath of the particularly vengeful Trump administration, especially as Hegseth has worked to eradicate civilian harm reduction processes to emphasize greater “lethality” as the military’s guiding principle.
这也促使行业领袖抛出一些看似反直觉的说法。OpenAI首席执行官萨姆·奥尔特曼等人一边警告其产品可能对人类生存构成危险,一边却向自封为“战争部长”的皮特·赫格塞斯推销该产品。AI Now Institute首席科学家、OpenAI前系统安全工程师海迪·赫拉夫告诉《拦截者》:“尽管人工智能公司不断警告人工智能可能有害,哪怕只是假设性的危害,但它们实际上却在推动将不可靠的产品接入国家安全攸关的基础设施,包括国防和核设施。鉴于人工智能在此类关键环境中的可靠性和准确性不足,这最终会导致灾难性事故。然而,如果承认这些问题,对人工智能公司而言就无利可图,也会拖慢其产品的普及速度。”美国人工智能企业或许仍清晰记得Anthropic近日的经历:该公司表示不愿让其软件被用于实施违宪的大规模国内监控或操控完全自主的武器,随后一度被政府列入黑名单,无法参与政府项目。若采取类似举措,限制这些公司的产品被用于从事间谍活动或杀人,可能会招致报复心极强的特朗普政府的愤怒,尤其是赫格塞斯一直致力于废除减少平民伤害的种种程序,并强调要将更强的“杀伤力”作为军方的指导原则。
The rapid adoption of their technology also represents a massive long-term financial boon to U.S. AI labs-turned-defense-contractors such as OpenAI, Google, and Anthropic — all of whom once vowed to not pursue military work because of potential harms. Despite these pledges, these firms now have access to the near-limitless Pentagon appropriations spigot, and the $200 million deals many of these firms signed with the Department last year pale in comparison with what the U.S. military could pay in the decades to come. Altman and his peers might also be less willing to sound the alarm on Pentagon-based AI threats because such a clarion call would be coming from inside the house.
他们的技术被迅速采用,也意味着OpenAI、谷歌和Anthropic这类由AI实验室转型而来的美国国防承包商将获得巨额长期财务收益——这些公司都曾誓言不涉足军事工作,原因是其可能造成伤害。尽管有此承诺,这些公司如今却能接触到五角大楼近乎无限的拨款渠道,而其中许多公司去年与国防部签署的2亿美元合同,与美国军方未来几十年可能支付的金额相比简直微不足道。奥特曼及其同行或许也不太愿意就五角大楼基于AI的威胁发出警报,因为这样的警钟将来自内部。
Earlier this month, The Intercept revealed contract documents detailing the intimate collaborative relationships between these companies and the military, in which the Silicon Valley giants would have a say in helping the Pentagon craft its overall military AI strategy. This closeness comes at a time when the Hegseth boasts of drone-striking civilian fishing boats across the Caribbean, and the commander-in-chief speaks casually of “annihilating” entire civilizations.
本月早些时候,《拦截》披露了合同文件,详细说明了这些公司与军方之间密切的合作关系,其中这些硅谷巨头将在帮助五角大楼制定其整体军事AI战略方面拥有发言权。这种亲密关系出现之际,赫格塞思正吹嘘在加勒比海对民用渔船实施无人机打击,而这位总司令还轻描淡写地谈论“消灭”整个文明。
Lucy Suchman, professor emerita of anthropology of science and technology at Lancaster University, attributed the difference in public alarm and media focus to the power that tech billionaires command when it comes to shaping narratives. Popular Silicon Valley ideologies, such as the Bay Area’s strain of “Rationalism” and the effective altruism movement, tend to consider a hypothetical “existential” threat from a superintelligence more worthy of attention than near-term violence. This priority, she said, that pervades the AI safety research community, obscures the real and present threat that today’s AI models pose in a world chockablock with nukes.
兰卡斯特大学科学技术人类学荣休教授露西·萨奇曼将公众警觉与媒体关注度的差异归因于科技亿万富翁在塑造叙事方面所掌握的权力。硅谷流行的意识形态,例如湾区的“理性主义”流派和有效利他主义运动,往往认为来自超级智能的假想“生存性”威胁比近期暴力更值得关注。她说,这种贯穿AI安全研究界的优先排序,掩盖了当今AI模型在一个核武器遍布的世界中所构成的真实而紧迫的威胁。
“Given the U.S. Dept. of War’s commitment to maximizing the speed and scale of target generation, AI-enabled targeting is already an existential threat for civilians on the ground, and when the targets are the military assets of nuclear armed states, the threat becomes global,” said Suchman, who has studied artificial intelligence since the 1970s and serves as a member of the nonprofit International Committee for Robot Arms Control. “The almost targeting of the Chinese ship might further underscore the danger not from rogue AI agents, but a rogue U.S. secretary of war.” Hegseth, who once derided the rules of engagement as “stupid,” extends this approach to artificial intelligence, too. Describing the Pentagon’s new AI strategy in a January memo that includes the word “accelerate” 24 different times, Hegseth’s laid out a plan with an intense focus on speed for speed’s sake. “Speed Wins,” Hegseth wrote. The war secretary explained the military would adopt experimental technologies as quickly as possible by “aggressively identifying and eliminating bureaucratic barriers to deeper Integration” of AI, without mention of ensuring that the technology actually works beforehand.
“鉴于美国国防部致力于加快目标识别的速度并扩大其规模,基于人工智能的目标识别系统已经对地面上的平民构成了生存威胁;而当这些目标涉及核国家的军事资产时,这种威胁就变成了全球性的威胁,”苏奇曼(Suchman)说道。他自20世纪70年代起就开始研究人工智能,并担任非营利组织“国际机器人武器控制委员会”(International Committee for Robot Arms Control)的成员。苏奇曼指出:“这次针对中国军舰的攻击事件可能进一步凸显出问题的根源并不在于‘失控的人工智能系统’,而在于‘行为失控的美国国防部长’。”赫格塞斯(Hegseth)曾嘲笑某些军事交战规则为“愚蠢的”,如今他将这种思维方式应用到了人工智能领域。在1月份发布的一份备忘录中,赫格塞斯详细描述了五角大楼的新人工智能战略;该备忘录中“加速”一词被重复使用了24次,整个战略的核心就是单纯追求速度。“‘速度就是胜利’,”赫格塞斯写道。这位国防部长表示,军方将尽快采用各种实验性技术,通过“积极消除官僚主义障碍”来推动人工智能技术的深入整合——却完全没有提及在采用这些技术之前是否需要确保它们能够真正发挥作用(即这些技术是否具备实际效用)。
“Given the U.S. Dept. of War’s commitment to maximizing the speed and scale of target generation, AI-enabled targeting is already an existential threat for civilians on the ground, and when the targets are the military assets of nuclear armed states, the threat becomes global,” said Suchman, who has studied artificial intelligence since the 1970s and serves as a member of the nonprofit International Committee for Robot Arms Control. “The almost targeting of the Chinese ship might further underscore the danger not from rogue AI agents, but a rogue U.S. secretary of war.” Hegseth, who once derided the rules of engagement as “stupid,” extends this approach to artificial intelligence, too. Describing the Pentagon’s new AI strategy in a January memo that includes the word “accelerate” 24 different times, Hegseth’s laid out a plan with an intense focus on speed for speed’s sake. “Speed Wins,” Hegseth wrote. The war secretary explained the military would adopt experimental technologies as quickly as possible by “aggressively identifying and eliminating bureaucratic barriers to deeper Integration” of AI, without mention of ensuring that the technology actually works beforehand. Heidi Kandiel, a legal adviser on new military technologies at the International Committee of the Red Cross, told The Intercept that speed itself carries the risk of military disaster for any country. “I think one of the central issues that comes up is speed and scale, which are often the touted benefits of AI,” she explained. “If a system is unreliable, is operating on poor quality or outdated data or is used within a flawed decision-making process, AI would allow these errors to be reproduced much faster and across a far greater number of targets or operations.”
海蒂·坎迪尔(Heidi Kandiel)是国际红十字会(International Committee of the Red Cross)负责研究新型军事技术的法律顾问,她告诉《The Intercept》:“速度本身就可能给任何国家带来军事灾难的风险。”她解释说:“我认为,速度和规模是人工智能(AI)常被吹捧的优点,但这两个因素实际上也可能带来严重问题。如果某个系统不可靠、使用的数据质量低下或已经过时,或者其应用过程存在缺陷,那么人工智能会加速这些错误的传播,使其影响到更多的目标或军事行动。”
Reining in a corporation is a very different prospect than putting fetters on a government. Policymakers who feel comfortable blasting the irresponsibility of Anthropic might lack the political will to levy similar criticisms against the president or the Pentagon. Madeline Berzak, a scholar at the University of Chicago’s Existential Risk Laboratory, told The Intercept that nuclear threats have long been met with paralysis (or a shrug) because people feel helpless in the face of them. You can tweet angrily at Dario Amodei, but it’s less satisfying to yell at U.S. Strategic Command. “This [China] incident and the lack of response is an attention problem that the nuclear field has been grappling with for decades,” Berzak said. “‘Doomsday messaging’ is what I call the nuke field’s default communications strategy, and I have this theory that it never works because the fear it tries to generate has nowhere to go.” “The threat of thermonuclear war is old news,” Suchman lamented.
与约束企业不同,约束政府则面临完全不同的挑战。那些敢于批评 Anthropic 公司不负责任行为的政策制定者,可能缺乏对总统或五角大楼(Pentagon)采取类似批评措施的勇气。芝加哥大学存在风险研究实验室(Existential Risk Laboratory)的学者玛德琳·贝尔扎克(Madeline Berzak)指出:“长期以来,面对核威胁时人们总是选择沉默或无动于衷,因为他们感到无能为力。你可以愤怒地发推文批评达里奥·阿莫代(Dario Amodei),但对着美国战略司令部(U.S. Strategic Command)大喊大叫却毫无意义。”贝尔扎克补充道:“这种‘中国事件’以及相关方缺乏有效应对的情况,其实反映了核领域多年来一直存在的问题——核武器相关信息的传播策略始终停留在‘末日警告’的水平上。我认为这种策略根本无效,因为人们所产生的恐惧感无处宣泄。”
Reining in a corporation is a very different prospect than putting fetters on a government. Policymakers who feel comfortable blasting the irresponsibility of Anthropic might lack the political will to levy similar criticisms against the president or the Pentagon. Madeline Berzak, a scholar at the University of Chicago’s Existential Risk Laboratory, told The Intercept that nuclear threats have long been met with paralysis (or a shrug) because people feel helpless in the face of them. You can tweet angrily at Dario Amodei, but it’s less satisfying to yell at U.S. Strategic Command. “This [China] incident and the lack of response is an attention problem that the nuclear field has been grappling with for decades,” Berzak said. “‘Doomsday messaging’ is what I call the nuke field’s default communications strategy, and I have this theory that it never works because the fear it tries to generate has nowhere to go.” “The threat of thermonuclear war is old news,” Suchman lamented. Some in the AI safety community acknowledge the long-standing threat of nuclear annihilation. They just consider it a lower priority than the as of yet unrealized rogue superintelligence. Daniel Kokotajlo left his role as an OpenAI researcher in 2024 over the company’s alleged recklessness and indifference toward risk. He is the co-author of AI 2027, a white paper laying out a hypothetical scenario in which superintelligent AI systems exterminate humankind with bioweapons to clear up space to build automated robotics factories. History is the reason Kokotajlo is less preoccupied by the prospect of an AI-triggered nuclear war. “It’s definitely something to be concerned about, and it’s something that ideally we would like to improve and fix. But it’s the sort of thing where, even if we do nothing, probably things will be fine for a decade or two or more.”
苏奇曼(Suchman)感叹道:“热核战争的威胁早已存在,但人们对此仍然习以为常、缺乏真正的警惕。”
Some in the AI safety community acknowledge the long-standing threat of nuclear annihilation. They just consider it a lower priority than the as of yet unrealized rogue superintelligence. Daniel Kokotajlo left his role as an OpenAI researcher in 2024 over the company’s alleged recklessness and indifference toward risk. He is the co-author of AI 2027, a white paper laying out a hypothetical scenario in which superintelligent AI systems exterminate humankind with bioweapons to clear up space to build automated robotics factories. History is the reason Kokotajlo is less preoccupied by the prospect of an AI-triggered nuclear war. “It’s definitely something to be concerned about, and it’s something that ideally we would like to improve and fix. But it’s the sort of thing where, even if we do nothing, probably things will be fine for a decade or two or more.”
AI安全社区中的一些人承认核毁灭的长期威胁。他们只是认为,与尚未实现的失控超级智能相比,这属于较低优先级。丹尼尔·科科塔伊洛因指责OpenAI鲁莽行事、漠视风险,于2024年离开了该公司研究员的职位。他是《AI 2027》的合著者,这份白皮书描绘了一个假设场景:超级智能AI系统利用生物武器灭绝人类,以腾出空间建设自动化机器人工厂。历史是科科塔伊洛不那么执着于AI引发核战争这一前景的原因。“这绝对值得关注,也是我们理想中希望改进和解决的问题。但这属于那种即使我们什么都不做,情况在未来十年、二十年甚至更久大概率也会没事的问题。”
To Kokotajlo and like-minded researchers, no matter how dangerous nuclear arsenals make the world, superintelligence is simply more dangerous than a bogus intelligence briefing on China. “It’s been 50 years and these sorts of crises don’t happen that often. So far, cooler heads have managed to prevail and figure out the truth in each case,” he said. “And so the probability that it goes all the way to nuclear war in the next five years is low. High enough to be a bit uncomfortable, but low overall. By contrast, if we do superintelligence in the next five years, I think the probability that it goes horribly wrong for us is high, not low.” Berzak, who leads her lab’s nuclear risk program, says this is the wrong calculation. She pointed to differing approaches to staving off the apocalypse embraced by the novel field of AI safety and nuclear caution that has existed since the dawn of the Cold War. “Attention often goes to what is new and interesting and not what is the most pressing risk,” she said.
在科科塔伊洛和志同道合的研究人员眼中,无论核武库让世界变得多么危险,超级智能都比一份关于中国的虚假情报简报更危险。“50年过去了,这类危机并不常发生。到目前为止,冷静的头脑总能占上风,在每次事件中弄清真相,”他说。“因此,未来五年演变为全面核战争的概率很低。虽高到让人有点不安,但总体上很低。相比之下,如果我们在未来五年实现超级智能,我认为它给我们带来灾难性后果的概率是高的,而非低的。”领导其实验室核风险项目的伯扎克表示,这种计算是错误的。她指出,新兴的AI安全领域与自冷战伊始就存在的核谨慎领域,在避免末日降临的方法上存在差异。“关注往往流向新颖有趣的事物,而非最迫在眉睫的风险,”她说。
To Kokotajlo and like-minded researchers, no matter how dangerous nuclear arsenals make the world, superintelligence is simply more dangerous than a bogus intelligence briefing on China. “It’s been 50 years and these sorts of crises don’t happen that often. So far, cooler heads have managed to prevail and figure out the truth in each case,” he said. “And so the probability that it goes all the way to nuclear war in the next five years is low. High enough to be a bit uncomfortable, but low overall. By contrast, if we do superintelligence in the next five years, I think the probability that it goes horribly wrong for us is high, not low.” Berzak, who leads her lab’s nuclear risk program, says this is the wrong calculation. She pointed to differing approaches to staving off the apocalypse embraced by the novel field of AI safety and nuclear caution that has existed since the dawn of the Cold War. “Attention often goes to what is new and interesting and not what is the most pressing risk,” she said.
冷静的头脑过去确实避免过核灾难,但人类会犯错,而且越来越容易产生一种诱惑:把我们的判断外包给被视为拥有超人类智慧的机器。伯扎克警告说:“仅仅因为如今出现了一个更为紧迫的新风险,而其后果与核战争相比尚未得到充分探索,并不意味着核战争不再构成风险;截至目前,我们基本上只是靠运气才避免了它。”
Cooler heads have indeed avoided nuclear disaster in the past, but human fallibility and the growing temptation to outsource our judgment to machines perceived as superhumanly brilliant carries an inherent danger, Berzak warned: “Just because there’s this new more pressing risk that has consequences that are rather underexplored compared to nuclear war, doesn’t mean that nuclear war is not a risk because we’ve so far avoided it basically with just luck.”