Nvidia launches safety platform to stop AI agents going rogue, says it could have prevented the Hugging Face hack

https://www.techspot.com/images2/news/ts3_thumbs/2026/09/2026-09-28-ts3_thumbs-269.jpg

Serving tech enthusiasts for over 25 years.
TechSpot means tech analysis and advice you can trust.

What just happened? Every day brings another story of an AI agent going rogue, hacking another organization, or generally doing something it's not supposed to do. Nvidia hopes to address this increasingly concerning trend with the launch of the Open Agent Safety Platform. Team Green has a lot of faith in the system: it says the platform could have prevented the infamous hack by OpenAI's agents on Hugging Face.

The platform combines OpenShell, Nvidia's open-source software for keeping agents within defined boundaries, with Sentry, a watchdog running on separate hardware.

The idea is to enforce restrictions outside the agent itself, putting security controls beyond the reach of software that might decide the rules are getting in its way, which is something that seems to be happening a lot recently.

OpenShell tracks agents' actions...

Copyright of this story solely belongs to www.techspot.com. To see the full text click HERE

Read more