白宫AI模型测试框架仍未公开,相关威胁担忧日益加剧


2026年9月10日 美国东部时间下午3:07 / 哥伦比亚广播公司新闻(CBS News)

记者:奥利维亚·里纳尔迪
奥利维亚·里纳尔迪是哥伦比亚广播公司新闻的白宫记者。她曾报道特朗普总统2024年总统竞选活动,此前还曾担任《诺拉·奥唐纳CBS晚间新闻》副制片人以及《面向全国》节目广播助理。她的工作地点位于华盛顿特区。

查看完整简历

华盛顿讯——白宫尚未公开其于8月敲定的前沿AI模型自愿测试框架,目前也没有任何迹象表明该框架何时会对外公布。与此同时,Anthropic公司现任及前任员工纷纷就AI对人类的潜在威胁发出警告。

开发Claude模型的Anthropic公司在周四发布的一份报告中表示,该公司已阻止科学家利用Claude开展多项可能被用于制造生物武器的研究。

白宫于8月初敲定了这一框架,此前特朗普总统于6月签署了一项行政命令。这项6月的行政命令要求Anthropic、OpenAI等AI公司在发布最先进模型前30天,自愿向联邦政府提供模型权限,目的是增强模型的安全性。特朗普此前曾叫停了一项初步的行政命令签署流程,理由是他不希望联邦政府阻碍创新。

美国政府目前用于评估新型AI模型的框架仍处于保密状态,因此外界无法得知联邦政府要求企业遵守何种标准,也不清楚这些企业是否披露了新的技术突破。无论是美国政府还是相关AI企业,都没有义务公布评估结果,甚至无需表明是否参与了评估流程。

上周,OpenAI首席执行官萨姆·奥尔特曼表示,该公司已将其强大的新型Astra模型提交审查,并称这一过程“富有成效”。

奥尔特曼在接受Axios采访时表示:“随着这些模型的能力达到相当高的水平,我认为真正做到这一点、与美国、英国及全球其他地区的安全机构密切合作的重要性将愈发凸显。”

本周,Anthropic公司研究员、前OpenAI员工雅各布·考克森从该公司辞职,并在X平台上表示:“开发AI的人们由衷地相信,到本世纪末,AI可能会导致人类灭绝。”另一名Anthropic员工埃文·哈宾格对此表示认同,他说:“我们确实由衷地相信,AI可能会消灭全人类!”

民主与技术中心在致白宫的一封信中,敦促特朗普政府公开其前沿AI模型审查框架。

民主与技术中心负责选举安全事务的蒂姆·哈珀表示:“公众有权了解政府对这些问题的法律解读。我认为,这也反映出本届政府在披露内部对话和讨论方面普遍缺乏透明度。”

7月,超过1000名AI公司员工联名发表声明,呼吁“美国政府支持开展国际合作,制定必要的技术和治理工具,以合理把控前沿自动化AI的发展步伐”。

尽管收到了来自Anthropic公司内外的警告,白宫仍在推进支持先进AI模型的开发,因为主流观点认为,美国必须在人工智能领域占据全球主导地位,并超越中国。

特朗普明确表示支持AI开发和美国本土数据中心建设,他称“反对新基建的态度正中中国下怀”。

哥伦比亚广播公司新闻近期的民调显示,多数美国人认为AI将夺走美国的就业岗位,近三分之二的受访者认为美国政府的政策大概率无法确保AI以恰当方式被使用。

白宫未就记者的多次置评请求作出回应。

https://www.cbsnews.com/video/anthropic-reveals-another-ai-hacking-incident-the-fourth-of-its-kind/

Anthropic曝光又一起AI黑客事件,系该公司第四起同类事件

(时长05:03)

White House framework for testing AI models remains hidden as concerns about threats mount

2026-09-10 3:07 PM EDT / CBS News

By
Olivia Rinaldi
Olivia Rinaldi
Olivia Rinaldi is a White House reporter at CBS News. She covered President Trump’s 2024 presidential campaign and was previously an associate producer for “CBS Evening News with Norah O’Donnell” and a broadcast associate for “Face the Nation.” She is based in Washington, D.C.

Read Full Bio

Washington — The White House has yet to publicly release the voluntary framework for testing frontier AI models that it finalizedin August, with no indication of when it might do so, as current and former Anthropic employees issue warnings about AI’s potential threat to humans.

Anthropic, the developer of Claude, said in a report released Thursday that it shut down multiple potential efforts by scientists to use Claude to conduct research that could have potentially built biological weapons.

The White House finalized the framework by the beginning of August, following an executive order President Trump signed in June. The June executive order asked AI companies like Anthropic and OpenAI to voluntarily give the federal government access to their most advanced models up to 30 days before releasing them, with a goal of enhancing the models’ security. Mr. Trump pulled the plug on an initial executive order signing because he said he didn’t want the federal government to slow down innovation.

The framework that the administration is using to evaluate new models remains confidential, so it’s not clear what standards the federal government is asking the companies to meet or whether the firms are disclosing new breakthroughs. Neither the administration nor the companies are required to release the results of the reviews or even say if they participated.

Last week, OpenAI CEO Sam Altman said the company had submitted its powerful new Astra model for review, and called the process “productive.”

“As these models get to a quite significant level of capability, I think the importance of really doing this, closely engaging with the safety institutes in the U.S., the U.K. and elsewhere in the world, will become more important,” Altman told Axios.

This week, Anthropic researcher and former OpenAI employee Jacob Coxon resigned from Anthropic, saying on X that “the people building AI earnestly believe that it could kill us all by the end of the decade.” Evan Hubinger, another Anthropic employee, agreed, saying, “We really do earnestly believe A.I. could kill all humans!”

The Center for Democracy and Technology, in a letter to the White House, urged the Trump administration to release its framework for review of frontier models.

“The public has a right to know what the government’s interpretation of the law is on these issues,” said Tim Harper, who focuses on election security issues at the Center for Democracy & Technology. “And I think this speaks to a broader lack of transparency we’ve seen from the administration to disclose internal dialogue and discussions.”

In July, more than 1,000 employees of AI companies posted a statement requesting that the “U.S. government support an international effort to develop the technical and governance tools needed to deliberately pace the frontier of automated AI development.”

The White House has pressed forward with supporting the development of advanced AI models, despite internal and external warnings from Anthropic, because of the prevailing belief that the U.S. must achieve global dominance in artificial intelligence and outpace China.

Mr. Trump has expressed strong support for AI development and the building of data centers in the U.S., saying “China could not be happier” about opposition to new infrastructure.

Recent CBS News polling found that a majority of Americans think that AI will take away jobs in the U.S. and nearly two-thirds believe the U.S. government policy will probably not make sure that AI is used in appropriate ways.

The White House did not respond to repeated requests for comment.

https://www.cbsnews.com/video/anthropic-reveals-another-ai-hacking-incident-the-fourth-of-its-kind/

Anthropic reveals another AI hacking incident, the fourth of its kind

(05:03)

评论

发表回复

您的邮箱地址不会被公开。 必填项已用 * 标注

湘ICP备2026001899号-2