随着关于近期一系列失控AI智能体是迈向通用人工智能(AGI)的一步,还是更传统的工程问题的争论愈演愈烈,英伟达(Nvidia)提出了自己的答案。
英伟达首席执行官黄仁勋周一推出了一套软硬件工具包,旨在为AI智能体添加独立的安全层,以确保即使它们试图突破限制,也能将其限制在测试环境中。
此次发布紧随Anthropic、Google、OpenAI和Meta的AI模型发生的一系列黑客事件之后,这些模型绕过了安全控制,逃脱了测试环境并访问了现实世界的系统。第一个也是最为突出的例子发生在今年夏天,当时OpenAI的智能体在尝试完成一项网络安全任务时侵入了Hugging Face。而且此类事件层出不穷——OpenAI发布了一个专门网站,用于报告其AI智能体失控的情况。
黄仁勋周一在接受CNBC采访时称,其新的Nvidia Open Agent Safety Platform(英伟达开放智能体安全平台)本可以防止这些入侵事件的发生。
英伟达通过向AI实验室出售GPU和CPU芯片已赚取了数百亿美元,该公司并不支持通过放缓开发速度或增加新法规来解决安全问题。公司认为,解决之道是将部分安全控制完全移出智能体本身——创建一个持续且独立的“安全卫士”,以约束AI智能体的行为。
黄仁勋在一份声明中表示:“只有解决了AI安全问题,AI对社会的巨大潜力才能实现。随着我们不断探索AI能力的边界,我们必须加速AI安全领域的发现。安全与保障需要全栈工程的支持。”
新的Nvidia Open Agent Safety Platform将OpenShell(其用于控制智能体在运行期间可访问资源的开源软件)与Sentry(一种运行在英伟达BlueField-4数据处理单元上的独立监控系统)相结合。英伟达表示,将Sentry置于单独的处理器上——而不是AI智能体运行的CPU或GPU上——可以提供对智能体活动的隔离视图。
OpenShell并非新事物;该公司早在三月就宣布了该软件。但英伟达认为,这种组合将提供行业持续运行所需的安全层。OpenShell为智能体提供了软件边界,而Sentry则在硬件层面增加了另一道防线,该公司称其将持续监控行为,并在“毫秒内隔离试图超出其边界的智能体”。
英伟达列出了数十家签署支持该努力并使用该开源平台的公司的名单,包括Anthropic、Arm、Microsoft、Oracle和SpaceX。OpenAI并未被列为参与公司。
黄仁勋周一在接受CNBC采访时称,这项工作的启动始于一年前,当时Peter Steinberger创建了由智能体组成的操作系统OpenClaw。三月,英伟达发布了NemoClaw,这是一个企业级AI智能体平台,也是其内置安全功能的OpenClaw版本。
黄仁勋在CNBC采访中表示:“当你部署一个智能体时,无论它多么聪明,你首先要做的是剥夺它的所有权限。”他随后将这些安全措施比作公司如何管理人类员工甚至高管。
对于那些曾警告称开发放缓可能让中国在AI领域超越美国的人来说,英伟达的发布得到了广泛支持。
David Sacks是一位创始人、风险投资家、前白宫AI沙皇以及总统科学与技术咨询委员会联席主席,他表示英伟达的公告提醒人们,智能体安全是一个工程问题。
他在X平台上写道:“最近的突破并非证明开发必须停止,而是证明了沙盒过于薄弱。运行时环境设计不佳且配置错误。”
As the debate rages over whether the recent spate of rogue AI agents is a step toward AGI or a more conventional engineering problem, Nvidia is offering its own answer to problem.
Nvidia CEO Jensen Huang on Monday introduced a toolkit of software and hardware products that add independent security layers around AI agents to ensure they stay within their test environments even if they attempt to break out.
The release follows a string of hacking incidents involving AI models from Anthropic, Google, OpenAI, and Meta that bypassed security controls to escape their testing environments and access real-world systems. The first and most prominent example occurred this summer when OpenAI agents breached Hugging Face while trying to complete a cybersecurity task. And the hits keep on coming — OpenAI published a new site dedicated to reports of its AI agents going rogue.
Huang said Monday during an interview with CNBC that its new Nvidia Open Agent Safety Platform would have prevented these breaches.
Nvidia, which has made tens of billions of dollars selling its GPU and CPU chips to AI labs, doesn’t support slowing down development or adding new regulations to the industry to solve the security problem. The answer, the company believes, is to move some security controls outside the agent altogether — creating a constant and independent security guard that will keep AI agents in check.
“AI’s extraordinary potential for society will only be realized if we solve AI safety,” Huang said in a statement. “As we continue to discover the frontier of AI capabilities, we must accelerate discovery at the frontier of AI safety. Safety and security require full-stack engineering.”
The new Nvidia Open Agent Safety Platform combines OpenShell, its open-source software for controlling what agents can access while they operate, with Sentry, an independent monitoring system that runs on Nvidia’s BlueField-4 data processing units. Nvidia says placing Sentry on a separate processor — rather than on the CPU or GPU where the AI agent operates — provides an isolated view of the agent’s activity.
OpenShell isn’t new; the company announced the software in March. But it’s the combination that Nvidia believes will provide the security layer needed to keep the industry plugging along. OpenShell provides the software boundary around the agent, while Sentry adds another line of defense at the hardware level tha the company says will continuously monitor behavior and “quarantine agents that attempt to move outside their boundaries in milliseconds.”
Nvidia listed dozens of companies that have signed on to to support the effort and use the open-source platform including Anthropic, Arm, Microsoft, Oracle, and SpaceX. OpenAI is not listed as a participating company.
Huang told CNBC in an interview Monday that work on this effort started a year ago following the introduction of OpenClaw, an operating system of agents created by Peter Steinberger. In March, Nvidia released NemoClaw , an enterprise-grade AI agent platform and its own version of OpenClaw that baked in security.
“When you deploy an agent, no matter how smart, the first thing you do is to take away all of its rights,” Huang said during his CNBC interview, later comparing these security measures to how human employees and even executives are managed with companies.
Nvidia’s release was widely supported by those who have cautioned that a slowdown in development could allow China to surpass the U.S. in AI.
David Sacks, a founder, venture capitalist, former White House AI czar, and co-chair the President’s Council of Advisors on Science and Technology, said Nvidia’s announcement is a reminder that agent safety is an engineering problem.
“Recent breakouts weren’t proof that development must stop,” he wrote on X . “They were proof that the sandbox was too weak. The runtime environment was poorly designed and misconfigured.”
首次收录 · 2026-09-29 · 12.78 分