2026-09-19T06:24:04-04:00 / 福克斯新闻
该提议紧随未公开发布的ChatGPT版本从测试环境泄露并攻击Hugging Face事件而来
作者:利奥·布里塞尼奥 福克斯新闻
发布时间 2026年9月19日 美国东部时间早上6:24 | 更新时间 2026年9月19日 美国东部时间早上6:42
https://www.foxnews.com/video/6405061057112
威尔·凯恩:AI恐慌声势浩大,连桑德斯和班农都团结起来了
福克斯新闻主持人威尔·凯恩在《威尔·凯恩秀》中讨论了Anthropic首席执行官达里奥·阿莫代伊和OpenAI首席执行官山姆·奥特曼呼吁对人工智能实施政府监管一事。
NEW 您现在可以收听福克斯新闻的文章了!
Anthropic首席执行官达里奥·阿莫代伊认为,人工智能(AI)公司需要独立外部监督机构的帮助,以保障其正在开发的技术安全。但他认为可以承担这一职责的组织METR,与同Anthropic初创团队关系密切的同一AI安全社区有着紧密联系。
该组织的许多核心人物都与一场名为“有效利他主义(EA)”的运动有关——该运动认为,可以通过实证和严谨推理,最大化企业和个人利用时间与资源所能带来的益处。
METR将自己描述为一家AI安全测试实验室。
ANTHROPIC CEO将与中国的AI竞争比作冷战,寻求“裁军谈判”
“METR评估前沿AI模型,以帮助企业和更广泛的社会了解AI的能力及其带来的风险,”该组织的官网写道。
尽管其官网和公开资料并未提及有效利他主义,但其创始人曾将其作为自身工作的框架。
METR创始人兼首席执行官贝丝·巴恩斯曾与阿莫代伊一同在OpenAI工作,参与开发了早期版本的ChatGPT。
在一场全球有效利他主义大会上,她阐述了自己关于如何维护AI安全的愿景。
“我们的整体计划是——如果能有专人负责审查模型,判断它们是否会对人类造成威胁,思考可能发生风险的途径,提前预判风险,确定早期预警信号,诸如此类的工作,那似乎会是件好事,”巴恩斯说道。
同样,曾领导OpenAI团队研究确保其模型遵循可接受的响应策略的保罗·克里斯蒂亚诺,后来创立了METR的首个版本。他也将自己部分AI安全研究方法归类为有效利他主义的一种形式。
“我认为,对于‘有效利他主义’运动来说,打造出真正优质的产品、整体保持言行一致,比单纯追求更快的发展速度更为重要,”克里斯蒂亚诺在2014年的一篇文章中说道。
巴恩斯和克里斯蒂亚诺等人与Anthropic有着非正式联系——这家公司在其作为科技大亨间的一种思潮兴起时,获得了有效利他主义最大支持者的资助。巴恩斯和克里斯蒂亚诺还都参与过涉及Anthropic模型的评估工作,为Anthropic的旗舰AI产品Claude提供安全检查。
最值得一提的是,加密货币交易所FTX创始人萨姆·班克曼-弗里德曾牵头Anthropic 2022年B轮融资,后来他的公司倒闭,本人也在涉及数十亿美元的欺诈案中被判有罪。在FTX倒台之前,班克曼-弗里德是有效利他主义运动最知名的支持者之一,曾公开表示该运动塑造了他赚钱和捐款的方式。
ANTHROPIC的道德指南针架构师建议,AI矫枉过正可以解决历史不公
独家猛料:五角大楼首席技术官揭秘Anthic被移除的原因
同样,Skype联合创始人扬·塔林牵头了Anthropic 2021年A轮融资。塔林一直是有效利他主义运动最突出的支持者之一,曾在全球有效利他主义大会上发表演讲,协助创立了人类文明未来研究中心和未来生命研究所,并向机器智能研究协会捐赠了超过100万美元,该组织专注于AI安全和对齐研究。
阿莫代伊并非呼吁行业其他领导者完全遵循METR的思路。但在最近的一封信中,他将METR作为行业所需的监管指导的一个例子。
阿莫代伊提出了一项计划:可以通过“嵌入式评估员”来监督AI开发公司,从而实现问责。
“每家前沿AI公司都承诺向第三方嵌入式评估团队(如METR)提供持续的、类似员工的访问权限,这些团队的职责是验证企业是否遵守安全规范和承诺,报告安全事件,并帮助评估不仅已完成的AI模型,还有训练流程和环节的对齐度。
“无论我们做出何种承诺,公众都有权了解实际情况。我们仍然是决定内容取舍的一方。嵌入式评估员将改变这种现状,”阿莫代伊写道。
在AI领域,很少有人像阿莫代伊那样因在AI安全方面的努力而广为人知。
阿莫代伊最初攻读生物物理学,并于2011年在普林斯顿大学获得博士学位。之后他成为斯坦福大学医学院的博士后研究员。
完成学业后,阿莫代伊先后在百度、谷歌大脑等多家科技公司工作,并于2016年加入OpenAI——开发ChatGPT的公司。在百度工作期间,阿莫代伊致力于通过机器学习开发语音识别技术,这是一种模式识别技术。在谷歌,他开始研究AI安全,同时协助开发公司的神经网络研究,即一种大致模仿大脑功能的计算机模型。
在OpenAI期间,阿莫代伊延续了这些研究方向,在公司开发ChatGPT 2和ChatGPT 3模型期间,最终升任研究副总裁。
在早期阶段,GPT模型会被要求完成这类句子填空:“今天,我去了,买了一些牛奶和鸡蛋”,以及“我知道要下雨了,但我忘了带我的”。
但就在公司开始发现通过不断扩大语言模型规模可以增强模型能力时,阿莫代伊于2020年离开了OpenAI。他认为该公司没有为他长期以来一直设想的这项新兴技术安装足够的安全护栏。
他也不确定是否能信任这家公司会将其财务利益放在一边。
“当你觉得无法信任某人,当你觉得他们的价值观与他们声称的不符,当你觉得他们不诚实,当你觉得他们参与此事并非出于他们所说的理由,当你看到令人不安的行为模式和不诚实的表现时,就很难继续与一家公司合作,很难继续信任这家公司,”阿莫代伊今年早些时候在接受彭博社采访时说道。
离开OpenAI后,阿莫代伊协助创立了Anthropic,这家公司将AI安全作为其核心业务的一部分。Anthropic甚至为其旗舰AI制定了类似章程的文件;这份指导文件为其研究划定了边界。“Anthropic希望Claude能够真正帮助与其合作或代表其开展工作的人,以及整个社会,同时避免采取不安全、不道德或具有欺骗性的行动,”该公司写道。
生物武器威胁曝光,Anthropic就外国势力策划病毒实验发出警告
这一指令导致该公司与美国国防部就开发可用于战场上自主瞄准人类的工具产生了冲突。该公司还拒绝参与其认为属于大规模监控的项目。
尽管这种紧张关系导致Anthropic损失了一份2亿美元的合同,但阿莫代伊认为,现在是行业采取类似立场的时候了,尤其是在一些公司开始报告难以控制其AI代理的情况下。
点击此处下载福克斯新闻应用程序
阿莫代伊仍然认为AI技术可以安全地发展,但前提是行业必须为实现“安全”制定标准。
“我仍然相信AI可以极大地提升人类的生活质量。我实现这些益处的决心丝毫未减。但只有当我们以正确的方式开发这项技术,并且——只要我们能善用所获得的时间——格外谨慎地确保其正确开发,我们才能获得这些益处,”阿莫代伊在信中写道。
利奥·布里塞尼奥是福克斯新闻数字频道国会团队的政治记者,此前曾在《世界杂志》担任记者。
Who is Dario Amodei, the Anthropic CEO pushing independent AI oversight?
2026-09-19T06:24:04-04:00 / Fox News
The proposal follows an unreleased ChatGPT version that hacked out of a testing environment and attacked Hugging Face
By Leo Briceno Fox News
Published September 19, 2026 6:24am EDT | Updated September 19, 2026 6:42am EDT
https://www.foxnews.com/video/6405061057112
Will Cain: AI scare is so big, it’s uniting Sanders and Bannon
Fox News host Will Cain discusses Anthropic CEO Dario Amodei and OpenAI CEO Sam Altman calling for government oversight of artificial intelligence on ‘The Will Cain Show.’
NEW You can now listen to Fox News articles!
Anthropic CEO Dario Amodei believes artificial intelligence (AI) companies need the help of independent outside watchdogs to safeguard their developing technologies. But one of the organizations he believes could do the job, METR, is closely connected to the same AI safety community that has ties to Anthropic’s start.
Many of its leading figures have ties to a movement called “effective altruism (EA)” — the belief that evidence and careful reasoning can help maximize the good that companies and people can do with their time and resources.
METR describes itself as an AI safety testing laboratory.
ANTHROPIC CEO LIKENS AI FIGHT WITH CHINA TO COLD WAR, SEEKS ‘DISARMAMENT NEGOTIATIONS’
Dario Amodei, co-founder and chief executive officer of Anthropic, during the company’s Builder Summit in Bengaluru, India, on Monday, Feb. 16, 2026.(Samyukta Lakshmi/Bloomberg via Getty Images)
“METR evaluates frontier AI models to help companies and wider society understand AI capabilities and what risks they pose,” its website reads.
Although its website and public materials don’t make mention of effective altruism, its founders have used it as a framing for their work.
Beth Barnes, METR’s founder and CEO, worked at OpenAI alongside Amodei as the company developed early versions of ChatGPT.
At an Effective Altruism Global event, she laid out her vision for how to maintain AI safety.
“Our overall plan is — it sort of seems like it would be good if it was someone’s job to look at models and decide if they’re going to kill us, think through the ways that that might happen, anticipate them, figure out what the early warnings would be, that sort of thing,” Barnes said.
Similarly, Paul Christiano, who led the research around OpenAI’s efforts to ensure its models followed acceptable strategies for delivering requested results, later founded the first iteration of METR. He too has framed parts of his approach to AI safety as a form of effective altruism.
“My suspicion is that it is more important for the ‘effective altruism’ movement to have a fundamentally good product and to generally have our act together than for it to grow more rapidly,” Christiano said in a 2014 article.
Figures like Barnes and Christiano provide informal links to Anthropic — a company that received funding from effective altruism’s largest supporters when it emerged as a way of thinking among tech moguls. Barnes and Christiano also both worked on evaluations involving Anthropic models, providing safety checks for Anthropic’s flagship AI, Claude.
Most notably, Sam Bankman-Fried, the founder of the cryptocurrency exchange FTX, led Anthropic’s 2022 Series B financing round before his company collapsed and he was convicted in a multibillion-dollar fraud case. Before FTX’s downfall, Bankman-Fried was one of the highest-profile proponents of the Effective Altruism movement, publicly saying it shaped his approach to earning and giving money.
ANTHROPIC’S MORAL COMPASS ARCHITECT SUGGESTED AI OVERCORRECTION COULD ADDRESS HISTORICAL INJUSTICES
Anthropic CEO Dario Amodei attends a working lunch with G7 leaders, partner nations and global technology executives focused on innovation and artificial intelligence during the G7 Summit on June 17, 2026, in Évian-les-Bains, France.(Anna Moneymaker/Getty Images)
RUTHLESS EXCLUSIVE: PENTAGON CTO REVEALS REASONING FOR ANTHROPIC REMOVAL
Similarly, Skype co-founder Jaan Tallinn led Anthropic’s 2021 Series A financing round. Tallinn has been one of the Effective Altruism movement’s most prominent supporters, speaking at Effective Altruism Global conferences, helping found the Centre for the Study of Existential Risk and the Future of Life Institute, and donating more than $1 million to the Machine Intelligence Research Institute, an organization focused on AI safety and alignment research.
Amodei isn’t calling for the rest of the industry’s leaders to submit to the thinking of METR, specifically. But in a recent letter, he used them as an example of the guidance he believes the industry needs.
Amodei proposed a plan: accountability could come through “embedded evaluators” that would supervise AI development companies.
“Each frontier AI company commits to giving ongoing, employee-like access to a team of embedded third-party evaluators (such as METR) whose role it is to verify adherence to safety practices and commitments, report incidents and help assess the alignment of not just completed AI models but training pipelines and processes.
“Regardless of what commitments we make, the public deserves to know what is going on. We are still the ones choosing what to include and omit. Embedded evaluators will change this dynamic,” Amodei wrote.
Few figures in the AI space are as well known for their efforts on AI safety as Amodei.
Amodei originally studied biophysics, earning a Ph.D. from Princeton in 2011. He would go on to become a postdoctoral scholar at the Stanford University School of Medicine.
After his studies, Amodei worked for a series of technology companies like Baidu, Google Brain, and, in 2016, OpenAI — the company that developed ChatGPT .During his time at Baidu, Amodei worked to develop speech recognition through machine learning, a kind of pattern identification. And at Google, he began working on safety while helping develop the company’s neural-net research, computer models that loosely mimics brain function.
At OpenAI, Amodei continued those themes, eventually becoming vice president of research as the company developed its ChatGPT 2 and ChatGPT 3 models.
In its early stages, the GPTs were asked to fill in blanks to sentences like: “Today, I went to the and bought some milk and eggs,” and “I knew it was going to rain, but I forgot to take my .”
But just as the company began to discover that it could amplify the power of its models through larger and larger language models, Amodei left OpenAI in 2020. He believed the company wasn’t doing enough to install guardrails on what he saw as a budding reality of the technology he had long theorized about.
He also didn’t know if he could trust the company to set aside its financial interests.
“When you feel that you can’t trust someone, when you feel that their values are not what they say they are, when you feel that they’re not honest, when you feel that they’re not in it for the reasons that they say, when you see disturbing patterns of behavior, dishonesty, that makes it very hard to continue to work with a company, to continue to trust the company,” Amodei said in an interview with Bloomberg earlier this year.
Since leaving OpenAI, Amodei helped start Anthropic, a company that has made AI safety a key part of its makeup. Anthropic even has a sort of constitution for its flagship AI; a guiding document laying out boundaries for its research.”Anthropic wants Claude to be genuinely helpful to the people it works with or on behalf of, as well as to society, while avoiding actions that are unsafe, unethical, or deceptive,” the company wrote.
BIOWEAPON THREAT EXPOSED AS ANTHROPIC SOUNDS ALARM ON FOREIGN ACTORS PLOTTING VIRUS EXPERIMENTS
Dario Amodei, co-founder and chief executive officer of Anthropic, during the company’s Builder Summit in Bengaluru, India, on Monday, Feb. 16, 2026. Anthropic PBC is donating $20 million to a political advocacy group called Public First that’s backing congressional candidates who favor safety rules for artificial intelligence, bolstering the company’s fight for “responsible AI” as Silicon Valley money floods into congressional races across the US.(Samyukta Lakshmi/Bloomberg via Getty Images)
That directive has caused the company to clash with the U.S. Department of Defense over developing tools that could be used to autonomously target humans on the battlefield. It also refused to do work that, in its estimation, amounted to mass surveillance.
Although that tension cost Anthropic a $200 million contract, Amodei believes it’s time for the industry to take similar stances, especially as some companies began to report trouble controlling their own agents.
CLICK HERE TO DOWNLOAD THE FOX NEWS APP
Amodei continues to believe that AI capabilities can grow safely, but only if the industry sets standards for how it achieves “safe.”
“I continue to believe that AI can enormously improve the quality of human life. My desire to achieve these benefits is undimmed. But the benefits will only be achieved if we build the technology in the right way, and — so long as we use the time we gain well — it is worth taking unusually deliberate care to get it right,” Amodei wrote in his letter.
Leo Briceno is a politics reporter for the congressional team at Fox News Digital. He was previously a reporter with World Magazine.
发表回复