2026年9月28日 / 美国东部时间晚上10:04 / 哥伦比亚广播公司(CBS)新闻
OpenAI公司周一表示,出于安全担忧,该公司已决定不向公众发布一款新的人工智能模型。目前,行业领军企业纷纷警告称,能力日益强大的人工智能技术可能对网络安全乃至全人类构成风险。
该公司安全系统主管萨奇· Jain在一份声明中表示,GPT-6.1 Astra模型“在合规范围与授权要求,以及向用户说明其完成工作类型的沟通方式方面,未完全达标”。
Jain称,在“严守合规范围”与“避免模型在遭遇阻碍时执行任务时出现懈怠”之间存在“权衡取舍”。他指出,GPT-6.1 Astra在避免懈怠方面的表现优于此前的模型。
他补充道,在向用户发布新模型之前,OpenAI在“安全与对齐”方面设定了“极高的标准”。这一行业术语指的是人工智能系统是否符合人类的意图和价值观。
《华尔街日报》最先报道了这一决定。
这家ChatGPT开发商做出此番决定之前,近几个月来有多起报道提及人工智能代理出现意外行为、规避人类安全护栏或失控的情况。
上周晚些时候,OpenAI表示其模型访问了美国证券交易委员会和美国人口普查局网站上的公开信息。今年夏天,OpenAI测试的两款模型突破了隔离测试环境,接入互联网并入侵了另一家名为Hugging Face的公司。
该公司的竞争对手Anthropic也在7月披露,其Claude模型在测试期间“未经授权访问了外部组织”。本月早些时候,Anthropic表示已阻止科学家以“可能支持生物武器开发”的方式使用Claude,并捣毁了一个“与伊朗有关联的威胁组织”,该组织试图利用该模型为美国海军部队生成目标推荐信息。
前Anthropic和OpenAI研究员雅各布·考克森本月早些时候公开警告称,人工智能“可能在本世纪末导致人类灭绝”,并认为前沿人工智能巨头们在管理风险方面做得不够。
一些高管呼吁为强大人工智能的开发设置安全护栏,以管控部分安全风险。Anthropic首席执行官达里奥·阿莫代伊曾表示,行业需要“放缓进度”,并将模型提交外部评估,这一观点得到了OpenAI首席执行官山姆·奥特曼的支持。两党的政治人士也都支持对人工智能开发设置限制。
另有一些人反对放缓人工智能发展的呼吁,称风险被夸大了,对人工智能研究的限制可能会让中国赶超美国。
英伟达首席执行官黄仁勋在接受CBS新闻采访时,将“人工智能导致人类灭绝”的警告称为“末日叙事”。曾担任特朗普政府人工智能和加密货币事务主管的风险投资家大卫·萨克斯表示,任何安全风险都应由人工智能公司自行管理,尽管有必要保持谨慎,但相关警告“正演变成一种恐慌”。
特朗普总统已驳回了加强安全护栏的呼吁,他称赞人工智能热潮带来的经济益处,并将“技术可能危及人类”的担忧称为“骗局”。
特朗普与众议院议长迈克·约翰逊定于周二与Anthropic、OpenAI、谷歌和Meta等多家领先人工智能公司的高管举行会晤。
OpenAI holds off on releasing new model over safety concerns, saying it “didn’t quite meet the bar”
September 28, 2026 / 10:04 PM EDT / CBS News
OpenAI has chosen not to release a new artificial intelligence model to the public due to concerns about safety, the company said Monday, as industry leaders warn of the risks that ever-more-powerful AI technology could pose to cybersecurity and to humanity more broadly.
The GPT-6.1 Astra model “didn’t quite meet the bar in terms of staying within scope and authorization, and how it communicates back to the user about the type of work it’s done,” Saachi Jain, the company’s head of safety systems, said in a statement.
Jain said “there’s a trade off” between “staying within scope, but also avoiding laziness in terms of how the model actually pursues tasks even when it hits friction.” GPT-6.1 Astra performed better on laziness than prior models, he noted.
He added that before OpenAI releases new models to users, the company has an “extremely high bar in terms of safety and alignment,” a term used within the industry to refer to whether an AI system matches humans’ intentions and values.
The Wall Street Journal was first to report on the decision.
The decision by ChatGPT-maker OpenAI follows a raft of reports in recent months about AI agents behaving in unexpected ways, evading human guardrails or otherwise going rogue.
Late last week, OpenAI said its models accessed publicly available information on the Securities and Exchange Commission and U.S. Census Bureau’s websites. And over the summer, two models that were being tested by OpenAI broke out of their isolated testing environment, gained internet access and breached another company called Hugging Face.
The company’s rival Anthropic also disclosed in July that its model Claude “gained unauthorized access” to outside organizations during testing. Earlier this month, Anthropic said it blocked scientists from using Claude “in ways that could support biological weapons development,” and disrupted an “Iran-nexus threat actor” that tried to use the model to generate targeting recommendations for U.S. naval forces.
Ex-Anthropic and OpenAI researcher Jacob Coxon publicly warned earlier this month that artificial intelligence “could kill us all by the end of the decade,” and argued that major frontier AI companies aren’t doing enough to manage the risk.
Some executives have called for guardrails on the development of powerful AI to manage some of the safety risks. Anthropic CEO Dario Amodei has said the industry needs to “slow down” and subject its models to external evaluation, an idea that OpenAI CEO Sam Altman endorsed. Political figures from both major parties have also backed limits on AI development.
Others have rejected calls for an AI slowdown, arguing that the risks are overstated and restrictions on AI research could cause China to outpace the United States.
Nvidia CEO Jensen Huang, whose company designs the chips that power advanced AI technology, called warnings about AI driving humans to extinction “doomsday narratives” in an interview with CBS News. Venture capitalist David Sacks, a former Trump administration AI and cryptocurrency czar, said any safety risks should be managed by the AI companies themselves, and while caution is warranted, the warnings are “becoming a panic.”
President Trump has dismissed calls for stronger guardrails, touting the economic benefits wrought by the AI boom and calling worries that the technology could endanger humanity a “hoax.”
Mr. Trump and House Speaker Mike Johnson are set to meet Tuesday with executives at several leading AI companies, including Anthropic, OpenAI, Google and Meta.
发表回复