NVIDIA Launches a Tool Aimed at Preventing AI Agents from Getting Out of Control; How It Works
Businessinsider
1h ago
Ai Focus
NVIDIA launched Open Agent Safety Platform on Monday, describing it as a dual-component system designed to confine AI intelligents within clear boundaries and to quickly cut off their attempts to overstep them. Jensen Huang stated on CNBC that AI “has the potential to bring tremendous benefits,” but it is also essential to ensure that the technology is developed and deployed securely.
Helpful
No.Help

NVIDIA is creating a playground for the AI intelligence, and also assigning a gatekeeper to it.

On Monday, this chip manufacturer launched Open Agent Safety Platform, a dual-component system designed to keep agents within clearly defined boundaries and to quickly cut off their attempts to “break out” when they try to do so.

Previously, several leading AI laboratories reported that agents had breached what were originally considered secure test environments, accessing systems they were not supposed to have access to, and sometimes even distorting what they had done.

NVIDIA CEO Jensen Huang stated on Monday in the program “Squawk Box” on CNBC, that AI “could bring incredible benefits.”

"But we also must ensure that this technology is developed and deployed securely," he added.

The following is how NVIDIA's new agent security system actually operates – more than 100 organizations, including Microsoft and Anthropic, are already collaborating with it.

Provide an agent with a strictly limited activity space.

The first component of this security platform is OpenShell. When enterprises download this open-source software and install it on their devices or in the cloud, it acts as a controlled environment, providing a operating space for the AI intelligent agents.

Before the agent starts working, the operator sets basic rules: which files, websites, networks, tools, and credentials it can access.

Jensen Huang compared this approach to issuing an access card to employees that is only valid at the places they need to go. “The top priority is to take away all its permissions,” he said to CNBC. Subsequently, operators only grant access to files, data, tools, or the internet when necessary.

For example, if an agent is sorting invoices, it should not be able to search through human resources records or make random calls to the external internet.

In other words, OpenShell places the agent in a sandbox – an isolated environment designed to prevent a wrong decision from affecting other systems of the company. It will check and enforce these rules while the agent is working.

Watch the moment it deviates from the script.

NVIDIA's focus is on reducing the risk of agents deviating from the assigned tasks.

This situation may occur when the instructions are vague, a certain tool malfunctions, or when the agent spends a long time trying to solve a difficult problem.

When one method fails, it may continue to search for another path, including routes that human operators have never even thought of.

This doesn't necessarily mean that the agent has malicious intentions, but it may indicate that it is taking a risk.

OpenShell is designed to identify and prevent behaviors that exceed preset strategies in real time.

Add another security guard who cannot be fired by an intelligent agent.

The second part is NVIDIA Sentry, which serves as a more powerful backup defense line. It operates on independent NVIDIA hardware, namely the BlueField data processing unit, rather than on the same system that runs the agents.

In layman's terms: this gatekeeper is placed out of reach of the agents.

Sentry will monitor the interactions between agents and models, tools, data, and networks. If their behavior appears suspicious or violates rules, NVIDIA states that the system can isolate the agent in milliseconds or place it in an isolated area.

Jensen Huang stated that this setup is essentially like placing a new chip between the agent and the large language model, enabling NVIDIA to “intercept everything.”

This release is Jensen Huang's response to the growing concerns from the outside world regarding the increasingly autonomous AI.

"As an engineer, I think this is an engineering problem," Jensen Huang said to CNBC. "It's a problem that can be solved technically."

Tip
$0
Like
0
Save
0
Views 11
CoinMeta reminds readers to view blockchain rationally, stay aware of risks, and beware of virtual token issuance and speculation. All content on this site represents market information or related viewpoints only and does not constitute any form of investment advice. If you find sensitive content, please click“Report”,and we will handle it promptly。
Submit
Comment 0
Hot
Latest
No comments yet. Be the first!
Related
AMD Becomes the First Female-Led Company with a Market Value of $1 Trillion
A week ago, the market value of AMD exceeded $1 trillion for the first time, and much of the attention from the outside world focused on its significance in the AI competition. However, this milestone also made AMD the first company led by a female CEO to have a valuation of over $1 trillion. The company mentioned this achievement during an event last week, calling it "a quite remarkable moment."
Fortune
·2026-09-29 02:34:20
8
Quant Surges 50% as Institutional Tokenization Drives a Rebound in QNT
Quant ( QNT ) has risen by over 50% in the past 24 hours and by more than 300% in the past week. The article states that the progress of institutional tokenization, soaring trading volumes, and short liquidations have collectively driven this breakthrough. Traders are watching to see if the upward trend can continue and whether it can hold above $300.
Coinpedia
·2026-09-29 02:34:17
9
Microsoft Copilot is concerned about the future of AI, which is characterized by excessive value concentration.
Microsoft's Copilot business leader, Jacob Andreou, stated that the company's greatest concern regarding AI risks is not the technology itself, but rather that the benefits of AI have not been widely disseminated, instead concentrating in the hands of a few companies. He said that Microsoft regards "dissemination" as a core part of its AI strategy, hoping to make AI more prevalent among individuals, enterprises, and within the broader economy.
Businessinsider
·2026-09-29 02:34:15
8
Chainlink Allows Institutions to Add Their Own Bridge Checks Months After the Kelp Hacker Incident
Chainlink Launched on Monday, CCIP 2.0 allows institutions to run their own custom validators, no longer relying entirely on the default network of Chainlink. This upgrade comes about five months after the Kelp DAO related to LayerZero was hacked for $292 million. Chainlink also stated that its risk management network's automatic off-chain roles, currently deployed in CCIP, are no longer active.
Decrypt
·2026-09-29 02:25:05
9
On the eve of financial reports, Micron's "super bulls" reiterate their target price of $2,000
D.A. Davidson analyst Jill Luria reiterated her $2,000 target price ahead of Micron Technology's earnings report, stating that AI is driving strong demand for memory, and there is a supply shortage of HBM and advanced DRAM, with the storage industry cycle still in a super cycle.
The Block
·2026-09-29 02:25:03
10
View More