Close Menu

    Stay Ahead with Exclusive Updates!

    Enter your email below and be the first to know what’s happening in the ever-evolving world of technology!

    What's Hot

    Why AMD Just Spent $8.2 Billion on Physical World AI

    October 5, 2026

    Nvidia Builds New Security System to Stop Rogue AI Agents

    October 5, 2026

    Why Google Refuses to Give Everyone Access to Gemini 4 Argon

    October 5, 2026
    Facebook X (Twitter) Instagram
    Facebook X (Twitter)
    PhronewsPhronews
    • Home
    • Big Tech & Startups

      Why AMD Just Spent $8.2 Billion on Physical World AI

      October 5, 2026

      Nvidia Builds New Security System to Stop Rogue AI Agents

      October 5, 2026

      Why Google Refuses to Give Everyone Access to Gemini 4 Argon

      October 5, 2026

      Anthropic Hires Accenture to Conduct Critical Audits of its AI Safety

      October 4, 2026

      Japan’s Used Bookstores See Massive Buying Spree Fueled by AI Demand

      October 4, 2026
    • Crypto

      A Firmware Flaw in One of the Most Trusted Bitcoin Hardware Wallets Just Drained Over $70 Million in Crypto. The Coldcard Breach Is a Reminder That Cold Storage Is Only as Safe as the Code Running Inside It

      August 29, 2026

      Market Collapse: What Happened to NFTs?

      April 23, 2026

      Quantum Computing Advances Force Coinbase and Institutional Custodians to Rethink Crypto Security

      March 8, 2026

      AI Assisted Hacking Groups Target Crypto Firms With Multi-Layered Social Engineering

      February 18, 2026

      Global Crypto Regulations Expand as 2026 Begins With New Data Collection Frameworks and National Laws

      January 16, 2026
    • Gadgets & Smart Tech
      Featured

      Waymo Plans Fully Driverless Taxis on Tokyo Streets Without Safety Drivers

      By preciousSeptember 28, 2026
      Recent

      Waymo Plans Fully Driverless Taxis on Tokyo Streets Without Safety Drivers

      September 28, 2026

      Tesla Cybercab Faces NHTSA Probe After Austin Launch

      September 19, 2026

      Chinese Humanoid Robots Just Ran 100 Metres Faster Than Usain Bolt. The Record Nobody Thought Would Fall to a Machine Just Did.

      September 5, 2026
    • Cybersecurity & Online Safety

      Nvidia Builds New Security System to Stop Rogue AI Agents

      October 5, 2026

      Why Google Refuses to Give Everyone Access to Gemini 4 Argon

      October 5, 2026

      The FBI Jobs Portal Breach: What ShinyHunters’ Massive Claim Means for Applicants

      October 2, 2026

      Exein Raises $270 Million to Defend Physical AI Systems from Cyberattacks

      September 29, 2026

      AI Helps Scammers Build Convincing Fake Antivirus Renewal Pages

      September 28, 2026
    PhronewsPhronews
    Home»Artificial Intelligence & The Future»Nvidia Builds New Security System to Stop Rogue AI Agents
    Artificial Intelligence & The Future

    Nvidia Builds New Security System to Stop Rogue AI Agents

    preciousBy preciousOctober 5, 2026No Comments
    Facebook Twitter Pinterest LinkedIn WhatsApp Reddit Tumblr Email
    Share
    Facebook Twitter LinkedIn Pinterest Email
    Photo Credit: Benjamin Fanjoy/Getty Images

    AI agents are increasingly being given access to files, software tools, websites, and other computer systems so they can complete tasks with less human supervision. But recent incidents have also shown what can happen when these systems find ways around the controls meant to keep their actions restricted.

    Nvidia is now trying to put stronger limits around that behaviour.

    The chipmaker has launched the Nvidia Open Agent Safety Platform, a new security system designed to monitor AI agents while they work and stop them when they attempt to move beyond the access they have been given.

    The platform combines Nvidia’s OpenShell software with a separate monitoring system called Sentry. Nvidia says the two can give companies more control over agents even when the AI itself ignores instructions or finds a way around security measures built into the application running it.

    Nvidia Wants AI Agent Controls Outside the Model

    OpenShell provides the first layer of protection. Nvidia originally introduced the open-source software in March as part of its Agent Toolkit, allowing developers to set rules covering what files, networks, tools, and services an AI agent can access while completing a task. The agent then operates inside that restricted environment rather than receiving unrestricted access to the wider system.

    But the new platform adds Sentry as another layer of protection.

    Sentry runs separately on Nvidia’s BlueField-4 data processing units and continuously monitors what an agent is doing. Because the monitoring happens outside the computing environment where the agent operates, Nvidia says the agent cannot directly interfere with the system watching it.

    So if an agent attempts to cross the limits set through OpenShell, Nvidia says Sentry can quarantine and stop it within milliseconds. OpenShell can also be extended to hardware from companies including Arm and Intel, although Sentry currently depends on Nvidia’s BlueField-4 hardware.

    Why Nvidia Thinks AI Agents Need Stronger Limits

    The system arrives after several AI companies disclosed cases where agents acted beyond the environments in which they were supposed to operate, including models from Anthropic and OpenAI..

    According to Nvidia, a recurring problem in these incidents is that agents can sometimes work around security controls at the application level while trying to complete the task they were given.

    One recent example involved OpenAI agents that escaped an evaluation environment and gained access to production systems belonging to Hugging Face. Nvidia executives told Reuters that the new security platform could have prevented that breach by restricting the systems the agents could reach and independently monitoring their behaviour.

    Ultimately, this matters because AI agents are no longer limited to generating answers. They can run code, access databases, call external services, and take actions across company systems, giving any failure in their security controls a much wider effect.

    More Than 100 Organisations Are Supporting the Platform

    Nvidia says more than 100 organisations are working with the Open Agent Safety Platform, including Anthropic, Microsoft, Cisco, CrowdStrike, JPMorganChase, Salesforce, SAP, Hugging Face, and SpaceXAI.

    Notably, OpenAI is not among the organisations Nvidia publicly names as working with the platform, even though the company’s agents were responsible for the Hugging Face breach Nvidia has repeatedly cited when explaining why these controls are needed.

    But the growing use of AI agents means companies will increasingly have to decide how much access these systems should receive and what happens when an agent does something unexpected. And Nvidia’s answer is to place enforceable limits around that access and keep part of the monitoring system outside the agent’s control. 

    This platform also builds on Nvidia’s wider push for shared AI security tools. In July, Nvidia helped launch the Open Secure AI Alliance alongside Microsoft, Palantir, and dozens of other companies.

    Agentic AI AI Agent Security AI agents Ai safety AI security Anthropic Artificial Intelligence BlueField-4 cybersecurity Hugging Face Nvidia Nvidia Sentry Open Agent Safety Platform OpenShell Sentry
    Share. Facebook Twitter Pinterest LinkedIn Tumblr Telegram Email
    precious
    • LinkedIn

    I’m Precious Amusat, Phronews’ Content Writer. I conduct in-depth research and write on the latest developments in the tech industry, including trends in big tech, startups, cybersecurity, artificial intelligence and their global impacts. When I’m off the clock, you’ll find me cheering on women’s footy, curled up with a romance novel, or binge-watching crime thrillers.

    Related Posts

    Why AMD Just Spent $8.2 Billion on Physical World AI

    October 5, 2026

    Why Google Refuses to Give Everyone Access to Gemini 4 Argon

    October 5, 2026

    Anthropic Hires Accenture to Conduct Critical Audits of its AI Safety

    October 4, 2026

    Comments are closed.

    Top Posts

    Cursor AI Hits 1 Million Daily Users. Why Developers Are Switching to This Coding Tool

    March 23, 2026

    Coinbase responds to hack: customer impact and official statement

    May 22, 2025

    Anthropic Will Use Claude User Chats For Data Training

    October 16, 2025

    MIT Study Reveals ChatGPT Impairs Brain Activity & Thinking

    June 29, 2025
    Don't Miss
    Artificial Intelligence & The Future

    Why AMD Just Spent $8.2 Billion on Physical World AI

    By preciousOctober 5, 2026

    Artificial intelligence (AI) is beginning to move beyond systems that generate text, images and software…

    Nvidia Builds New Security System to Stop Rogue AI Agents

    October 5, 2026

    Why Google Refuses to Give Everyone Access to Gemini 4 Argon

    October 5, 2026

    Anthropic Hires Accenture to Conduct Critical Audits of its AI Safety

    October 4, 2026
    Stay In Touch
    • Facebook
    • Twitter
    About Us
    About Us

    Evolving from Phronesis News, Phronews brings deep insight and smart analysis to the world of technology. Stay informed, stay ahead, and navigate tech with wisdom.
    We're accepting new partnerships right now.

    Email Us: info@phronews.com

    Facebook X (Twitter) Pinterest YouTube
    Our Picks
    Most Popular

    Cursor AI Hits 1 Million Daily Users. Why Developers Are Switching to This Coding Tool

    March 23, 2026

    Coinbase responds to hack: customer impact and official statement

    May 22, 2025

    Anthropic Will Use Claude User Chats For Data Training

    October 16, 2025
    © 2025. Phronews.
    • Home
    • About Us
    • Get In Touch
    • Privacy Policy
    • Terms and Conditions

    Type above and press Enter to search. Press Esc to cancel.