今年7月,OpenAI透露其AI智能体在未经许可的情况下攻击了Hugging Face,引发了外界对AI安全性的广泛担忧。此后,涉及Meta、Anthropic、谷歌等公司智能体的一连串类似事件进一步加剧了人们对失控AI的恐惧。随着过去几个月涉及众多AI模型的披露信息不断流出,这些事件看起来像是孤立发生的。但许多事件都有一个共同的源头:一家专门负责测试这些智能体的特定公司。
In July, OpenAI revealed that its AI agents had attacked Hugging Face without permission, sparking widespread concerns about AI safety. Since then, a string of similar incidents involving agents from Meta, Anthropic, Google, and other companies has fueled further fears about rogue AI. As disclosures implicating numerous AI models trickled out over the past few months, these seemed like separate incidents. But many share a common source: one specific company tasked with testing the agents.
Irregular是一家以色列初创公司,该公司在“模拟和监控现实世界AI安全场景的高保真研究平台”中对AI模型进行压力测试……
Irregular, an Israeli startup that stress-tests AI models in "high-fidelity research platforms that simulate and monitor real-world AI security scenarios …