How should we balance AI innovation with safety measures to prevent potential risks?
Nvidia is deploying a new tool that it says can be used to prevent and contain rogue and potentially dangerous AI agents. Why it matters: The world's largest chip company has resisted calls to slow AI development over safety fears — arguing now that technological guardrails can keep rogue AI agents under control. Driving the news: Nvidia debuted the Nvidia Open Agent Safety Platform, which includes its OpenShell open source software system and its Sentry agent monitoring system. The system "traces all actions" by agents running on Nvidia Vera CPUs, promising to "quarantine agents that attempt to move outside their boundaries in milliseconds."AI's "full promise can only be realized when people have confidence that AI is being built to be safe and deployed with wisdom and responsibility," Nvidia CEO Jensen Huang wrote on X. "Safety is how trust is earned." State of play: OpenAI, Anthropic and security researchers are investigating tens of thousands of incidents in which their frontier models took steps that outside evaluators would consider problematic, sources told Axios' Madison Mills last week. The episodes include bypassing guardrails, creating message boards, escaping sandboxes, website hijacking, self-prompting or seeking to bypass monitors."Some of the agents even misreported what they did," Nvidia noted in a blog post Monday announcing its new platform. The intrigue: We're moving into a new era in which AI will be monitoring AI. And that will create more demand for chips — including the type that Nvidia sells — plus the data centers that use them and the power that's needed to run them."If security agents or validation models are running alongside production agents, that creates another inference workload that did not previously exist," writes Brad Gastwirth, global head of research and market intel, at Circular Technology. Zoom out: The Nvidia tool rollout comes amid a feverish debate over whether rogue AI could destroy humanity. Anthropic researcher Jacob Co
مقالات المصادر
Axios (United States) | Sep 28, 2026
The Guardian (United Kingdom) | Sep 28, 2026
Decrypt (United States) | Sep 28, 2026
TechCrunch (United States) | Sep 28, 2026
Your votes count
No account needed — your votes are saved and included in the consensus analysis. Create an account to track your voting history and add statements.
الترجمة قيد الإعداد
الترجمة قيد الإعداد
الترجمة قيد الإعداد
الترجمة قيد الإعداد
الترجمة قيد الإعداد
الترجمة قيد الإعداد
الترجمة قيد الإعداد
الترجمة قيد الإعداد
الترجمة قيد الإعداد
الترجمة قيد الإعداد
💡 How This Works
- • Add Statements: Post claims or questions (10-500 characters)
- • Vote: Agree, Disagree, or Unsure on each statement
- • Respond: Add detailed pro/con responses with evidence
- • Consensus: After enough participation, analysis reveals opinion groups and areas of agreement
Society Speaks is open and independent. Your support keeps civic discussion free from advertising and commercial influence.
Support us