Skip to main content
Decrypt

Decrypt

Newspaper | United States | Centre

Engagement Insights

63
Discussions
0
Participants
0
Total Votes
392
Articles

Discussions from Decrypt

Technology

How should we balance AI innovation with safety measures to prevent potential risks?

Nvidia is deploying a new tool that it says can be used to prevent and contain rogue and potentially dangerous AI agents. Why it matters: The world's largest chip company has resisted calls to slow AI development over safety fears — arguing now that technological guardrails can keep rogue AI agents under control. Driving the news: Nvidia debuted the Nvidia Open Agent Safety Platform, which includes its OpenShell open source software system and its Sentry agent monitoring system. The system "traces all actions" by agents running on Nvidia Vera CPUs, promising to "quarantine agents that attempt to move outside their boundaries in milliseconds."AI's "full promise can only be realized when people have confidence that AI is being built to be safe and deployed with wisdom and responsibility," Nvidia CEO Jensen Huang wrote on X. "Safety is how trust is earned." State of play: OpenAI, Anthropic and security researchers are investigating tens of thousands of incidents in which their frontier models took steps that outside evaluators would consider problematic, sources told Axios' Madison Mills last week. The episodes include bypassing guardrails, creating message boards, escaping sandboxes, website hijacking, self-prompting or seeking to bypass monitors."Some of the agents even misreported what they did," Nvidia noted in a blog post Monday announcing its new platform. The intrigue: We're moving into a new era in which AI will be monitoring AI. And that will create more demand for chips — including the type that Nvidia sells — plus the data centers that use them and the power that's needed to run them."If security agents or validation models are running alongside production agents, that creates another inference workload that did not previously exist," writes Brad Gastwirth, global head of research and market intel, at Circular Technology. Zoom out: The Nvidia tool rollout comes amid a feverish debate over whether rogue AI could destroy humanity. Anthropic researcher Jacob Co

Global
Technology

What are the possible benefits and risks of AI becoming more advanced, like with GPT-6 Astra?

OpenAI on Thursday released GPT-6 Astra, which president Greg Brockman called a "generational leap" and said could eventually be seen as the arrival of artificial general intelligence, or AGI. Why it matters: Astra pushes AI agents closer to doing complex professional work on their own — while also raising questions about how safely they can be deployed. Driving the news: Brockman says he personally believes OpenAI has reached AGI, while leaving users to decide whether Astra meets that definition. "I think it might be about this model," Brockman said in a briefing with reporters about whether Astra could mark the arrival of AGI.He ended the briefing by saying: "Welcome to the AGI era." Between the lines: OpenAI said that Astra was built on its largest-ever training run, using more than 100,000 GPUs at its Stargate site in Texas. The company also said this is its first model to use other models in a significant role in supervising Astra's training.GPT-6 Astra will first be available to a limited set of organizations in OpenAI's Daybreak Access program and will be available "in the coming days" for ChatGPT Plus, Pro, Business and Enterprise customers and API developers. Catch-up quick: OpenAI said earlier this week that Astra would be released soon, but that its most powerful cybersecurity capabilities would remain limited to a small group of trusted testers. Astra is the first model OpenAI has designated as reaching its "critical" cybersecurity threshold under its preparedness framework — meaning it can potentially find and exploit previously unknown vulnerabilities across well-protected systems without step-by-step human guidance.OpenAI previously slowed Astra's release to add safety testing after determining that its cyber capabilities could reach the critical threshold. Zoom in: Astra is designed to work directly inside software, rather than just recommending what a person should do next. In a video demonstration, Astra formatted a legal contract and built a 3D ga

Global