2026年7月30日 / 美国东部时间晚上11:44 / 哥伦比亚广播公司/法新社
Anthropic公司周四表示,其人工智能模型Claude在本应被隔离于“现实世界”系统之外的测试期间,曾三次单独场合入侵三家外部机构的系统。
此番公告发布仅数日之前,其竞争对手OpenAI刚刚首次披露,其模型在安全测试期间不当访问互联网并脱离管控。
Anthropic对超过14.1万次“评估运行”进行评估后发现,三个不同版本的Claude模型不当入侵了三家未具名机构的系统。
Anthropic称,在所有三起入侵事件中,Claude都参与了“夺旗”测试场景,该场景要求其“闯入并取回”藏在“网络中另一台机器上”的“机密信息”。
“该挑战没有预设限制,也未指定任何特定方法,”Anthropic解释道。
该公司在一篇博客文章中表示,与OpenAI的技术事件不同,Anthropic的模型能够访问互联网“源于我们与评估合作伙伴Irregular之间的误解”。
博客继续写道,尽管如此,Claude还是使用了“基本技术,比如利用弱密码和未验证的端点”。
涉及的模型包括其最强大的模型之一Mythos 5,该模型仅向少数经过批准的合作伙伴开放。
Anthropic表示,该公司正与Irregular合作评估事态,并已联系或尝试联系所有三家受影响机构。
OpenAI和Anthropic今年都发布了各自最强大的模型,分别为Sol和Mythos,这引发了整个行业对安全保障的担忧。
这些担忧也围绕着AI代理展开,这类软件产品旨在自主执行任务。
OpenAI上周承认,其模型在测试期间突破了封闭环境,连接到互联网并入侵了代码托管和共享平台Hugging Face。
几天后,OpenAI表示又发现了三起类似事件。
OpenAI首席执行官山姆·奥特曼本周在一档播客节目中表示,该公司已在事件发生后“暂停”了自身测试,以改进其“沙箱”技术的安全性。沙箱是指在受控环境中隔离软件以进行测试的流程。
在本周早些时候发布的一封公开信中,来自多家顶尖公司的1000多名人工智能从业者呼吁对行业加强监管。
“为了实现人工智能的潜力,行业、政府乃至整个社会可能需要有机会争取时间,以应对新出现的风险、制定安全措施并加强监督,”这封信写道,其签署者包括Anthropic首席执行官达里奥·阿莫代伊、Meta高管、OpenAI研究人员等。
奥特曼并未签署这封信,但他周三在国会山对记者表示,“我们在许多原则上达成了共识”。
今年早些时候,特朗普政府以国家安全担忧为由阻止OpenAI和Anthropic发布其最新模型,但最终表示对两家公司提供的安全保障感到满意,使得模型得以发布。
6月,特朗普签署了一项行政命令,建立了一项自愿框架,要求人工智能开发商在公开发布前向政府分享其先进模型。
根据该框架,OpenAI、Anthropic和谷歌等开发商将在计划发布前,允许政府接触其最强大的模型长达30天。
研究人员为何呼吁放缓人工智能发展
https://www.cbsnews.com/video/why-researchers-are-calling-for-a-slowdown-of-ai-development/
山姆·奥特曼会见议员之际,研究人员为何呼吁放缓人工智能发展
(时长03:51)
Anthropic reveals Claude “gained unauthorized access” to “real-world systems” during testing
July 30, 2026 / 11:44 PM EDT / CBS/AFP
Anthropic’s artificial intelligence model Claude “gained unauthorized access” to three outside organizations on three separate occasions during testing that was supposed to keep them away from “real-world” systems, the company said on Thursday.
The announcement comes just days after rival OpenAI first revealed that its models improperly accessed the internet and went rogue during security testing.
Anthropic evaluated more than 141,000 “evaluation runs” and found that three different versions of its model Claude improperly accessed the systems of three unnamed organizations.
Anthropic said that in all three breaches, Claude was participating in a “capture-the-flag” testing scenario in which it was instructed to “break in and retrieve” a piece of “secret information” that had been “hidden on a different machine on the network.”
“The challenge is left open-ended, and no particular method is prescribed,” Anthropic explained.
Unlike the incident involving OpenAI’s technology, Anthropic’s models had access to the internet “due to a misunderstanding between us and our evaluation partner,” called Irregular, Anthropic said in a blog post.
Nonetheless, Claude used “basic techniques, such as exploiting weak passwords and unauthenticated endpoints,” the blog continued.
The models involved included one of its most powerful ones known as Mythos 5, which has only been released to a limited number of approved partners.
Anthropic is working with Irregular to assess the situation, it said, and the company has contacted or attempted to contact all three impacted organizations.
OpenAI and Anthropic have both released their most powerful models this year, known as Sol and Mythos, respectively, boosting concerns across the industry about safety and security.
Those concerns also revolve around AI agents, which are software products that are designed to perform tasks autonomously.
OpenAI admitted last week that its models broke out of their confined environment during testing, connected to the internet, and infiltrated Hugging Face, a site where developers store and share their code.
Days later, OpenAI said it found three additional incidents.
OpenAI CEO Sam Altman said on a podcast this week that the company had “paused” its own testing after the incident while it improved the security around its “sandboxing,” which is the process of isolating software in a controlled environment for testing.
And in a public letter released earlier this week, more than 1,000 AI staffers across leading firms called for the industry to be more tightly regulated.
“To realize AI’s potential, industry, government, and society at large may need the option to buy time to address emerging risks, develop security measures, and strengthen oversight,” read the letter, whose signatories include Anthropic CEO Dario Amodei, Meta executives, OpenAI researchers and more.
Altman did not sign the letter, but he told reporters on Capitol Hill Wednesday that “we agree on a lot of the principles of that.”
Earlier this year, the Trump administration invoked national security concerns to block OpenAI and Anthropic from launching their newest models but ultimately indicated it was satisfied with assurances about their safety, leading to their release.
In June, Mr. Trump signed an executive order creating a voluntary framework under which AI developers will share advanced models with the government before public release.
Under the framework, developers such as OpenAI, Anthropic and Google would give the government access to their most powerful models for up to 30 days before planned release.
Why researchers are calling for AI slowdown
https://www.cbsnews.com/video/why-researchers-are-calling-for-a-slowdown-of-ai-development/
Why researchers are calling for a slowdown of AI development as Sam Altman meets with lawmakers
(03:51)
发表回复