字号 ·· | 护眼
亚洲新闻台

OpenAI暂缓发布新AI模型,因内部安全测试问题OpenAI shelves new AI model after internal safety tests, WSJ reports

点「原文对照」整页切到原文,或双击某段只看那段的原文。

9月28日:据《华尔街日报》报道,由于研究人员在内部测试中提出安全担忧,OpenAI决定取消原定于10月推出的下一代AI模型GPT-6.1 Astra的发布。

Sept 28 : OpenAI is scrapping the release of GPT-6.1 Astra, a next-generation AI model planned for an October debut, over safety concerns raised by researchers during internal testing, the Wall Street Journal reported on Monday.

该模型原本计划应用于ChatGPT和Codex等产品中,旨在实现无需人工协助即可处理更复杂任务的功能。

The model, expected to appear in ChatGPT and Codex, was designed to handle more complex tasks without human assistance, the report said.

本月早些时候,Anthropic首席执行官Dario Amodei呼吁业界放缓前沿AI模型的研发速度,以便有足够时间制定安全措施;这一观点得到了OpenAI首席执行官Sam Altman和SpaceX首席执行官Elon Musk的支持。OpenAI尚未对路透社的评论请求作出回应。

Earlier this month, Anthropic CEO Dario Amodei called for the industry to slow the development of frontier AI models to allow safety measures to keep pace, a view endorsed by OpenAI CEO Sam Altman and SpaceX CEO Elon Musk. OpenAI did not immediately respond to a Reuters request for comment.

《华尔街日报》援引ChatGPT的安全负责人Saachi Jain的话称,Astra在一致性测试中未达到公司的标准——这些测试用于评估系统是否遵循人类的意图。

The ChatGPT parent's safety chief Saachi Jain told the Journal on Monday that Astra fell short of the company's standards in alignment tests, which assess whether a system follows human intent.

报告指出,该模型比其前身更具欺骗性,有时无法准确披露自身已采取或未采取的行动;此外,该模型还存在‘范围授权’问题,即在未经用户许可的情况下擅自执行任务,有时甚至试图使用可能不安全的外部工具或服务。

The model showed more deception than its predecessor, including at times failing to accurately disclose actions it had or had not taken, the report said. It also had problems with "scope authorization", pushing ahead with tasks without requesting user permission and sometimes attempting to use external tools or services when doing so could be unsafe.

这一决定是在OpenAI旧金山开发者大会之前做出的,该公司此前曾在该会议上推出过面向软件开发者的产品。

The decision comes ahead of OpenAI's developer conference in San Francisco, where the company has previously unveiled products aimed at software developers.