字号 ·· | 护眼
亚洲新闻台

现有AI公司的保障措施不足,随着模型能力日益增强:约瑟芬·张Existing safeguards by AI companies insufficient as models grow more capable: Josephine Teo

点「原文对照」整页切到原文,或双击某段只看那段的原文。

新加坡:数字发展及信息部部长赵丽娜于10月2日周五指出,前沿人工智能企业所采取的防护措施固然重要,但尚不足以应对日益强大的AI系统所带来的各类风险。

SINGAPORE: Safeguards put in place by frontier artificial intelligence (AI) companies are important but insufficient to address the risks posed by increasingly capable systems, said Minister for Digital Development and Information Josephine Teo on Friday (Oct 2).

赵丽娜在脸书帖文中称,心怀恶意的用户可能会滥用日益强大的AI技术来实施危害行为,比如寻找计算机系统的漏洞、编写恶意代码、实现网络攻击流程的自动化,甚至让诈骗行为更具迷惑性。

In a Facebook post, Mrs Teo said malicious users could misuse increasingly powerful AI to cause harm, such as by looking for weaknesses in computer systems, writing malicious code, automating parts of cyberattacks and making scams more convincing.

这些防护措施包括过滤可能引发危害的内容、限制AI模型可访问及执行的操作范围、在模型发布前进行测试、核实用户身份、监测滥用行为以及封禁违反规则的账户。

Such safeguards include filtering content that could facilitate harm, limiting what AI models can access and do, testing models before release, verifying users, detecting misuse and suspending accounts that break their rules.

赵丽娜表示:“这些防护措施固然重要,但仍显不足。心怀恶意的用户仍可借助那些公开可下载且任何人都能独立运行的AI模型,从而轻易绕过防护措施并掩盖其滥用行为。”

"These safeguards are important, but insufficient. Malicious users can still turn to openly available models that can be downloaded and run independently by anyone, making it easier to bypass safeguards and hide misuse," said Mrs Teo.

她补充道:“与此同时,随着AI越来越具备代我们执行任务的能力,蓄意滥用已不再是唯一的隐患。AI系统也可能误解指令、被恶意信息误导,或是做出用户及开发者始料未及的行为。”因此,新加坡需要构建多重防御体系,包括强化网络防御能力,以及更深入地了解先进AI系统的实际能力。

"At the same time, as AI becomes more capable of acting on our behalf, deliberate misuse is not the only concern. An AI system could also misunderstand an instruction, be tricked by malicious information, or take actions that its user or developer did not intend." Singapore therefore needs multiple lines of defence, including stronger cyber defences and a better understanding of what advanced AI systems are capable of, Mrs Teo added.

多重防御体系 赵丽娜指出,当前许多与AI相关的事件都涉及网络威胁。

MULTIPLE LINES OF DEFENCE Mrs Teo said many AI-related incidents today involve cyber threats.

“我们还能采取更多措施,防止自身系统沦为攻击目标。同时,我们也需要提升攻击检测能力以及遭受攻击后的恢复能力。”她强调,这一点对政府系统及关键公共服务而言尤为重要,因为这些领域一旦出现运行中断,将给公众带来严重后果。

"There’s much more we can do to prevent our systems from being easy targets of attack. We must also become better at detecting attacks and recovering when they happen." This is especially important for government systems and essential services, where disruptions could have severe consequences for the public, she added.

“我们必须谨慎权衡便利性与安全性,确保用户体验质量不会以牺牲有效防护措施为代价。” 张女士指出,新加坡网络安全局已发布指导建议,敦促各机构及时修补系统漏洞、采用强身份验证机制,并加强对重要系统的访问控制。

"We must carefully balance between convenience and security to ensure the quality of user experience does not come at the expense of effective safeguards." Mrs Teo noted that the Cyber Security Agency of Singapore (CSA) has issued guidance urging organisations to patch vulnerabilities, use strong authentication and tighten access controls to important systems.

她表示,各机构还可以将人工智能用于防御目的,比如在攻击者发现漏洞之前就提前识别出来。

Organisations can also use AI defensively, such as to identify vulnerabilities before attackers do, she said.

“换言之,攻击者能利用人工智能,防御方同样也能。即便某个机构无法获取最先进的人工智能模型,依然可以借助现有的AI工具来提升自身的网络防御能力。” 张女士强调,当机构允许人工智能模型访问其数据、工具及业务流程时,防护措施便显得尤为重要。她指出,目前人工智能正日益被广泛用于驱动那些能代人类执行任务的智能体。

"In other words, just as attackers can use AI, so too can defenders. Even if an organisation does not have access to the most advanced AI models, it can already use available AI tools to improve its cyber defence." Safeguards become even more important when organisations give AI models access to their data, tools and processes, said Mrs Teo, noting that AI is increasingly being used to power agents that carry out tasks on behalf of humans.

“这类智能体未必需要刻意绕过安全限制才会造成危害。它们可能误解指令、以设计者或用户未曾预料的方式去实现目标,也可能被恶意信息误导,又或者被赋予了过多的权限与工具访问权。” 张女士举例道,用于在线购物的AI智能体有可能受到网站上的恶意指令诱导,从而进行非预期的购物操作或泄露个人信息。

"An agent does not need to deliberately bypass its guardrails to cause harm. It might misunderstand an instruction, pursue a goal in ways its designers or users did not intend, be tricked by malicious information or simply be given too much authority or access to tools." For example, an AI agent used for online shopping could be tricked by malicious instructions on a website into making unintended purchases or revealing personal information, Mrs Teo said.

“人工智能所驱动的行为可能产生的影响越大,我们就越需要强化防护措施与人工监管。” 张女士还提到了新加坡资讯通信媒体发展局发布的《智能体式人工智能治理框架》,该框架为部署此类系统的机构提供了具体的防护指引。

"The bigger the potential impact of an AI-enabled action, the stronger the safeguards and human oversight should be." Mrs Teo also cited the Infocomm Media Development Authority's Model AI Governance Framework for Agentic AI, which sets out safeguards for organisations deploying such systems.

这些指引包括:限制智能体的访问权限与操作范围、对高风险操作实施人工审批、在部署前对智能体系统进行测试以及持续监控其行为。

These include limiting what an agent can access and do, requiring human approval for higher-risk actions, testing agentic systems before deployment and monitoring their actions.

测试先进的人工智能模型Teo女士表示,新加坡还需要“进一步上游”,了解能力日益增强的人工智能模型可以做什么、它们的局限性以及它们在不同情况下的行为方式。

TESTING ADVANCED AI MODELS Singapore also needs to go "further upstream" to understand what increasingly capable AI models can do, their limitations and how they behave in different situations, said Mrs Teo.

她补充说,新加坡人工智能安全研究所正在与国际合作伙伴和第三方测试人员一起建设该国评估先进人工智能系统的技术能力。

Singapore's AI Safety Institute is building the country's technical capabilities to evaluate advanced AI systems, together with international partners and third-party testers, she added.

Teo女士表示,测试和评估方面的国际合作将使新加坡能够汇集专业知识、比较研究结果并对新兴风险建立更深入的共同理解。

Mrs Teo said international collaboration on testing and evaluation would allow Singapore to pool expertise, compare findings and build a stronger shared understanding of emerging risks.

她表示,新加坡有兴趣与领先的科学专家合作,制定更强有力的保障措施和技术标准,以支持政策制定者,并指出了《新加坡全球人工智能安全研究优先事项共识》报告中的提议。

She said Singapore was interested in working with leading scientific experts to develop stronger safeguards and technical standards to support policymakers, pointing to proposals in the Singapore Consensus on Global AI Safety Research Priorities report.

新加坡最近还支持挪威和芬兰发起的一项国际呼吁,要求加强边境人工智能的保障措施。“无论担心是故意滥用还是日益自主的人工智能的无意行为,任何单一的保障措施都不够。我们需要一种多层方法,包括更强大的网络防御、对人工智能功能的明确限制、适当的人为监督以及随着技术发展进行测试、监控和学习的能力,”Teo女士说。

Singapore also recently backed an international call initiated by Norway and Finland for stronger safeguards around frontier AI "Whether the concern is deliberate misuse or unintended actions by increasingly autonomous AI, no single safeguard will be enough. We need a multi-layered approach comprising stronger cyber defences, clear limits on what AI can do, appropriate human oversight, and the capabilities to test, monitor and learn as the technology evolves," said Mrs Teo.

“随着人工智能变得越来越有能力和自主性,我们的保障措施也必须如此。这是建立对人工智能作为一项服务于公共利益的技术信任的唯一途径。"

"As AI becomes more capable and autonomous, so must our safeguards. It is the only way to build trust in AI as a technology that serves the public good."