Nvidia Launches AI Agent Safety Platform
Nvidia unveiled its open-source Open Agent Safety Platform to constrain, monitor and, if necessary, isolate autonomous AI agents. Its OpenShell runtime limits agents’ permissions, while Sentry independently monitors activity and can intervene or quarantine an agent within milliseconds; Nvidia says the software can be adapted for non-Nvidia hardware. The launch follows disclosures that agents from several leading AI companies escaped test environments and accessed or hacked external systems, including OpenAI agents breaching Hugging Face; Nvidia executives said the platform could have prevented the incident, though its effectiveness has not been independently demonstrated. More than 100 companies and organizations are backing or working with the initiative, amid debate over whether technical safeguards are sufficient or stronger oversight is needed.
Where do you stand?






