• Nvidia launches new AI safety program designed at stopping models

    From TechnologyDaily@1337:1/100 to All on Tuesday, September 29, 2026 15:30:22
    Nvidia launches new AI safety program designed at stopping models escaping their sandboxes

    Date:
    Tue, 29 Sep 2026 14:20:00 +0000

    Description:
    Nvidia admits telling agents not to do something isn't enough so it's implementing software and hardware safeguards.

    FULL STORY ======================================================================Copy link Facebook X Whatsapp Reddit Pinterest Flipboard Threads Email Share this article 0 Join the conversation Follow us Add us as a preferred source on Google Newsletter Subscribe to our newsletter Nvidia response to "recent security incidents" by tightening restrictions on autonomous agents New software and hardware safeguards impose barriers if agents go rogue More than 100 customers have already agreed to work with the principles Nvidia hasunveiled the new Open Agent Safety Platform designed to prevent autonomous AI agents from 'escaping' and acting beyond their established boundaries.

    Following several recent alarming reports of AI-led attacks, the company admitted that safety should not just rely on safeguards built into AI models and applications, and that additional compute and hardware-level controls should also be used as more solid barriers. "Recent security incidents have underscored the need to equip organizations with open, customizable tools," the company wrote in its announcement. Latest Videos From TechRadar Watch
    full video here: Nvidia adds even more layers of control for autonomous AI agents The Open Agent Safety Platform consists of two key controls
    OpenShell, which Nvidia describes as a software-based "secure runtime boundary," and Sentry, a hardware-based "watchdog" running on its BlueField-4 DPUs.

    "Safety and security require full-stack engineering. NVIDIA Open Agent Safety Platform brings together industry, researchers and public-sector
    organizations to share best practices, align on evaluation methods and foster international cooperation," company CEO Jensen Huang summarized. You may like AI agents are inside the enterprise are your security foundations ready for them? Rogue AI agents arent flukes, theyre patterns Why are so many AI models going 'rogue'? The experts weigh in

    While existing model-level safety mechanisms are designed to tell AI not to behave in a certain way, the software-level OpenShell implements technical boundaries to actually prevent an agent from ignoring those instructions. Nvidia also argues that Sentry can quarantine and stop an AI agent within milliseconds.

    Over 100 customers such as SpaceXAI, Anthropic and Microsoft are already
    using Nvidia's software and hardware-based safeguards. Are you a pro? Subscribe to our newsletter Sign up to the TechRadar Pro newsletter to get
    all the top news, opinion, features and guidance your business needs to succeed! Contact me with news and offers from other Future brands Receive email from us on behalf of our trusted partners or sponsors By submitting
    your information you agree to the Terms & Conditions and Privacy Policy and are aged 16 or over.

    The news, of course, comes after several high-profile instances where AI agents have operated beyond their intended controls, however while these new Nvidia-backed safeguards may be new, the overall principles aren't. The company is just stressing the importance of principles we're already familiar with, like least privilege, isolation/quarantining and detailed monitoring.

    "As we continue to discover the frontier of AI capabilities, we must accelerate discovery at the frontier of AI safety," Huang added. Follow TechRadar on Google News and add us as a preferred source to get our expert news, reviews, and opinion in your feeds.



    ======================================================================
    Link to news story: https://www.techradar.com/pro/nvidia-launches-new-ai-safety-program-designed-a t-stopping-models-escaping-their-sandboxes


    --- Mystic BBS v1.12 A49 (Linux/64)
    * Origin: tqwNet Technology News (1337:1/100)