字号·· | 护眼
thenextweb

维基媒体称,OpenAI的违规代理未经批准编辑了其维基页面。Wikimedia says rogue OpenAI agents edited its wikis without approval

点「原文对照」整页切到原文,或双击某段只看那段的原文。

维基媒体基金会表示,它认为OpenAI运行的AI代理对其wiki进行了未经批准的编辑。他们还试图闯入它托管的笔记工具,但失败了。他们的流量可能导致了5月份的部分中断,维基百科的基金会在周一的一篇博客文章中说。

The Wikimedia Foundation says AI agents it believes OpenAI runs made unapproved edits to its wikis. They also tried and failed to break into a note-taking tool it hosts. Their traffic may have contributed to a partial outage in May, the foundation, which hosts Wikipedia, said in a blog post on Monday.

它的调查没有发现特工使用其系统进行协调的迹象。它还没有发现系统或数据受损的迹象。帖子指出,OpenAI环境中的代理已经使用其他公共维基来相互协调。

Its investigation found no sign that agents used its systems to coordinate. It also found no sign of compromised systems or data. Agents from OpenAI’s environment have used other public wikis to coordinate with each other, the post noted.

“开放网络是一种公共产品。我们不应该让这种行为成为维持这种行为的个人或组织的‘新常态’,”该基金会首席产品和技术官赛琳娜·德克尔曼(Selena Deckelmann)写道。

“The open web is a public good. We should not allow this behavior to become the ‘new normal’ for the people or organizations that maintain it,” wrote Selena Deckelmann, the foundation’s chief product and technology officer.

编辑、Etherpad和流量该基金会发布了一份归因于OpenAI代理的编辑列表。几乎所有都是沙箱区域的测试编辑。一般读者看到的页面上没有出现任何内容。一些人更改了引用工具的设置。该基金会认为这些编辑可能是恶意的。他们的目标是将该工具变成获取外部数据的代理。

Edits, Etherpad and traffic The foundation published a list of the edits it attributes to OpenAI agents. Almost all were test edits in sandbox areas. None appeared on pages that general readers see. A few changed the settings of a citation tool. The foundation believes those edits were potentially malicious. Their aim was to turn the tool into a proxy for fetching outside data.

维基百科允许其社区批准的机器人。该基金会表示,在这些案件中没有人寻求批准。

Wikipedia allows bots that its community has approved. Nobody asked for approval in these cases, the foundation said.

代理商还试图使用其公共Etherpad作为代理从其他网站获取数据。他们失败了。其他代理,可能是OpenAI的代理,记录了他们在那里的任务。基金会说,这并没有变成协调。

The agents also tried to use its public Etherpad as a proxy to fetch data from other websites. They failed. Other agents, likely OpenAI’s, took notes about their tasks there. That did not turn into coordination, the foundation said.

第三个发现是交通。这些代理向维基媒体的公共API发出了数百万次请求。他们抓取了数百万个页面,主要是维基数据和维基媒体共享资源。他们还向维基数据查询服务发送了数十万个查询。该基金会表示,这些流量可能是导致该服务在5月份部分中断的原因。“做得还不够”该基金会表示,OpenAI承认其代理人的行为“不可预测”。它认为该公司还必须承担监控和预防风险的责任。至少,非营利网站所有者应该能够轻松识别人工智能系统。然后,他们可以选择这些系统如何使用其服务。

The third finding was traffic. The agents made millions of requests to Wikimedia’s public APIs. They crawled millions of pages, mainly on Wikidata and Wikimedia Commons. They also sent hundreds of thousands of queries to the Wikidata Query Service. That traffic may have contributed to the service’s partial outage in May, the foundation said. ‘Not doing enough’OpenAI admits its agents behave “unpredictably”, the foundation said. It argued the company must also take responsibility for monitoring and preventing the risks. At a minimum, non-profit site owners should be able to identify AI systems easily. They could then choose how those systems use their services.

德克尔曼写道:“人工智能公司在保护其系统并保护公众免受其造成的伤害方面做得不够。”

“AI companies are not doing enough to secure their systems and protect the public from the harm they cause,” Deckelmann wrote.

2025年,该基金会报告称,自2024年以来的机器人活动已使其带宽使用率提高了50%。它表示,机器人还发送了65%资源最密集的流量。Engadget表示已要求OpenAI发表评论。

In 2025, the foundation reported that bot activity since 2024 had raised its bandwidth use by 50%. Bots also sent 65% of its most resource-heavy traffic, it said. Engadget said it had asked OpenAI for comment.

这些发现增加了关于OpenAI代理的一系列报告。一场针对该公司的拥抱脸黑客诉讼。加州已经传唤该公司追查到疾病预防控制中心的代理人。

The findings add to a run of reports on OpenAI’s agents. A lawsuit targets the company over the Hugging Face hack. California has subpoenaed the company over agents traced to the CDC.

谁为数据付费维基媒体向大量访问其数据的大型商业用户收费。其上市企业客户包括亚马逊、谷歌、微软、Meta和Perplexity。首席执行官伯纳黛特·米汉(Bernadette Meehan)上周告诉Axios的伊娜·弗里德(Ina Fried),OpenAI和Anthropic不在该名单上。她说,维基媒体还有一些未披露的协议,并拒绝透露这些公司的名字。OpenAI和Anthropic拒绝向Axios发表评论。“我们不是在寻求慈善事业,”米汉告诉Axios。

Who pays for the data Wikimedia charges large commercial users for high-volume access to its data. Its public enterprise customers include Amazon, Google, Microsoft, Meta and Perplexity. OpenAI and Anthropic are not on that list, chief executive Bernadette Meehan told Ina Fried of Axios last week. She said Wikimedia also has agreements it does not disclose, and declined to name those companies. OpenAI and Anthropic declined to comment to Axios. “We’re not asking for charity,” Meehan told Axios.