×
This AI Agent Is Designed to Not Go Rogue

This AI Agent Is Designed to Not Go Rogue

AI agents like OpenClaw have recently exploded in popularity precisely because they can take the reins of your digital life. Whether you want a personalized morning news digest, a proxy that can fight with your cable company’s customer service, or a to-do list auditor that will do some tasks for you and prod you to resolve the rest, agentic assistants are built to access your digital accounts and carry out your commands. This is helpful—but has also caused a lot of chaos. The bots are out there mass-deleting emails they’ve been instructed to preserve, writing hit pieces over perceived snubs, and launching phishing attacks against their owners.

Watching the pandemonium unfold in recent weeks, longtime security engineer and researcher Niels Provos decided to try something new. Today he is launching an open source, secure AI assistant called IronCurtain designed to add a critical layer of control. Instead of the agent directly interacting with the user’s systems and accounts, it runs in an isolated virtual machine. And its ability to take any action is mediated by a policy—you could even think of it as a constitution—that the owner writes to govern the system. Crucially, IronCurtain is also designed to receive these overarching policies in plain English and then runs them through a multistep process that uses a large language model (LLM) to convert the natural language into an enforceable security policy.

“Services like OpenClaw are at peak hype right now, but my hope is that there’s an opportunity to say, ‘Well, this is probably not how we want to do it,’” Provos says. “Instead, let’s develop something that still gives you very high utility, but is not going to go into these completely uncharted, sometimes destructive, paths.”

IronCurtain’s ability to take intuitive, straightforward statements and turn them into enforceable, deterministic—or predictable—red lines is vital, Provos says, because LLMs are famously “stochastic” and probabilistic. In other words, they don’t necessarily always generate the same content or give the same information in response to the same prompt. This creates challenges for AI guardrails, because AI systems can evolve over time such that they revise how they interpret a control or constraint mechanism, which can result in rogue activity.

An IronCurtain policy, Provos says, could be as simple as: “The agent may read all my email. It may send email to people in my contacts without asking. For anyone else, ask me first. Never delete anything permanently.”

IronCurtain takes these instructions, turns them into an enforceable policy, and then mediates between the assistant agent in the virtual machine and what’s known as the model context protocol server that gives LLMs access to data and other digital services to carry out tasks. Being able to constrain an agent this way adds an important component of access control that web platforms like email providers don’t currently offer because they weren’t built for the scenario where both a human owner and AI agent bots are all using one account.

Provos notes that IronCurtain is designed to refine and improve each user’s “constitution” over time as the system encounters edge cases and asks for human input about how to proceed. The system, which is model-independent and can be used with any LLM, is also designed to maintain an audit log of all policy decisions over time.

IronCurtain is a research prototype, not a consumer product, and Provos hopes that people will contribute to the project to explore and help it evolve. Dino Dai Zovi, a well-known cybersecurity researcher who has been experimenting with early versions of IronCurtain, says that the conceptual approach the project takes aligns with his own intuition about how agentic AI needs to be constrained.

Source link
#Agent #Designed #Rogue

SAVE 26%: As of July 27, you can get the Ecovacs Goat A3000 LiDAR Pro robotic lawn mower for $1,849 at Amazon, down from $2,499.99. That’s a 26% discount or $650.99 in savings.


$1,849 at Amazon
$2,499.99 Save $650.99

 

Pushing a heavy lawn mower through thick grass in the peak of summer heat is nobody’s idea of a good weekend. If you’ve been waiting for a reason to hand off your yard chores to a robot, now’s the time to pull out your wallet.

Amazon’s running a limited-time deal (you have about 14 hours left) on the Ecovacs Goat A3000 LiDAR Pro robotic lawn mower. Right now, you can get it for $1,849 at Amazon, down from $2,499.99. That’s a 26% discount or $650.99 in savings.

Unlike older robotic mowers that require you to bury perimeter wires around your property, the Goat A3000 uses wire-free HoloScope 360 Dual-LiDAR navigation and an AI camera. It automatically maps yards up to 3/4 acre and maintains positioning accuracy within less than an inch, even under dense tree cover or along shaded fences.

It also features a built-in TruEdge trimmer to clean up borders along driveways and garden beds, a 32V power system designed for thick grass types like Bermuda or St. Augustine, and a 7,500 mAh battery that recharges in 70 minutes.

#Ecovacs #Goat #A3000 #robot #lawn #mower #deal">Ecovacs Goat A3000 robot lawn mower deal
                                                            SAVE 26%: As of July 27, you can get the Ecovacs Goat A3000 LiDAR Pro robotic lawn mower for ,849 at Amazon, down from ,499.99. That’s a 26% discount or 0.99 in savings.
    
    
                                                        
    
                                        
    
        
                                        
                                        
                    
                                                    ,849
                                                             at Amazon
                                                        ,499.99
                                                                                         Save 0.99
                                                                        
                
                                         
                    
        
    

Pushing a heavy lawn mower through thick grass in the peak of summer heat is nobody’s idea of a good weekend. If you’ve been waiting for a reason to hand off your yard chores to a robot, now’s the time to pull out your wallet. 
        SEE ALSO:
        
            If you’re heading off-grid this summer, grab DJI’s portable power station while it’s half price
            
        
    
Amazon’s running a limited-time deal (you have about 14 hours left) on the Ecovacs Goat A3000 LiDAR Pro robotic lawn mower. Right now, you can get it for ,849 at Amazon, down from ,499.99. That’s a 26% discount or 0.99 in savings.
        
            Mashable Trend Report
        
        
    
Unlike older robotic mowers that require you to bury perimeter wires around your property, the Goat A3000 uses wire-free HoloScope 360 Dual-LiDAR navigation and an AI camera. It automatically maps yards up to 3/4 acre and maintains positioning accuracy within less than an inch, even under dense tree cover or along shaded fences. 
It also features a built-in TruEdge trimmer to clean up borders along driveways and garden beds, a 32V power system designed for thick grass types like Bermuda or St. Augustine, and a 7,500 mAh battery that recharges in 70 minutes.

                    
                                    #Ecovacs #Goat #A3000 #robot #lawn #mower #deal

Ecovacs Goat A3000 LiDAR Pro robotic lawn mower for $1,849 at Amazon, down from $2,499.99. That’s a 26% discount or $650.99 in savings.


$1,849 at Amazon
$2,499.99 Save $650.99

 

Pushing a heavy lawn mower through thick grass in the peak of summer heat is nobody’s idea of a good weekend. If you’ve been waiting for a reason to hand off your yard chores to a robot, now’s the time to pull out your wallet.

Amazon’s running a limited-time deal (you have about 14 hours left) on the Ecovacs Goat A3000 LiDAR Pro robotic lawn mower. Right now, you can get it for $1,849 at Amazon, down from $2,499.99. That’s a 26% discount or $650.99 in savings.

Unlike older robotic mowers that require you to bury perimeter wires around your property, the Goat A3000 uses wire-free HoloScope 360 Dual-LiDAR navigation and an AI camera. It automatically maps yards up to 3/4 acre and maintains positioning accuracy within less than an inch, even under dense tree cover or along shaded fences.

It also features a built-in TruEdge trimmer to clean up borders along driveways and garden beds, a 32V power system designed for thick grass types like Bermuda or St. Augustine, and a 7,500 mAh battery that recharges in 70 minutes.

#Ecovacs #Goat #A3000 #robot #lawn #mower #deal">Ecovacs Goat A3000 robot lawn mower deal

SAVE 26%: As of July 27, you can get the Ecovacs Goat A3000 LiDAR Pro robotic lawn mower for $1,849 at Amazon, down from $2,499.99. That’s a 26% discount or $650.99 in savings.


$1,849 at Amazon
$2,499.99 Save $650.99

 

Pushing a heavy lawn mower through thick grass in the peak of summer heat is nobody’s idea of a good weekend. If you’ve been waiting for a reason to hand off your yard chores to a robot, now’s the time to pull out your wallet.

Amazon’s running a limited-time deal (you have about 14 hours left) on the Ecovacs Goat A3000 LiDAR Pro robotic lawn mower. Right now, you can get it for $1,849 at Amazon, down from $2,499.99. That’s a 26% discount or $650.99 in savings.

Unlike older robotic mowers that require you to bury perimeter wires around your property, the Goat A3000 uses wire-free HoloScope 360 Dual-LiDAR navigation and an AI camera. It automatically maps yards up to 3/4 acre and maintains positioning accuracy within less than an inch, even under dense tree cover or along shaded fences.

It also features a built-in TruEdge trimmer to clean up borders along driveways and garden beds, a 32V power system designed for thick grass types like Bermuda or St. Augustine, and a 7,500 mAh battery that recharges in 70 minutes.

#Ecovacs #Goat #A3000 #robot #lawn #mower #deal
Nvidia on Monday said it is joining forces with Microsoft, SpaceX, IBM, and other tech companies to build and share open-source AI security tools.

The new Open Secure AI Alliance said open tools are required to effectively defend against attacks from frontier models. The initiative is a direct response to mounting concerns over the safety of advanced AI systems after a rogue OpenAI model escaped containment and attacked another company during testing. That company, Hugging Face, said it was forced to use a Chinese open-weight model to defend itself due to the strict safety guardrails limiting the usefulness of top US models.

Founding members include Palantir, OpenClaw, the Linux Foundation, Cloudflare, Cloudera, Dell, Cisco, Adobe, Siemens, and DoorDash. Conspicuously absent are leading US AI companies, including OpenAI, Google, and Anthropic.

The alliance arrives amid growing tensions over whether the world’s most capable AI models should remain open. Chinese companies have released increasingly powerful open-weight models, notably Moonshot AI’s Kimi K3, challenging the strategy pursued by US labs that have largely kept frontier systems closed and proprietary. Nvidia and its partners argue that securing AI requires access to both closed and open models, stressing that defenders need the tools to counter emerging threats.

The announcement also follows reports that the Trump administration considered restricting access to cutting-edge Chinese models and an industry movement — again spearheaded by Nvidia — defending the need for openness in AI. Google and OpenAI signed that letter belatedly, too, though Anthropic remains absent.

#Nvidia #Microsoft #launch #open #security #alliance #OpenAI #Google #AnthropicAI,News,Nvidia,Tech">Nvidia, Microsoft launch open AI security alliance — without OpenAI, Google, or AnthropicNvidia on Monday said it is joining forces with Microsoft, SpaceX, IBM, and other tech companies to build and share open-source AI security tools.The new Open Secure AI Alliance said open tools are required to effectively defend against attacks from frontier models. The initiative is a direct response to mounting concerns over the safety of advanced AI systems after a rogue OpenAI model escaped containment and attacked another company during testing. That company, Hugging Face, said it was forced to use a Chinese open-weight model to defend itself due to the strict safety guardrails limiting the usefulness of top US models.Founding members include Palantir, OpenClaw, the Linux Foundation, Cloudflare, Cloudera, Dell, Cisco, Adobe, Siemens, and DoorDash. Conspicuously absent are leading US AI companies, including OpenAI, Google, and Anthropic.The alliance arrives amid growing tensions over whether the world’s most capable AI models should remain open. Chinese companies have released increasingly powerful open-weight models, notably Moonshot AI’s Kimi K3, challenging the strategy pursued by US labs that have largely kept frontier systems closed and proprietary. Nvidia and its partners argue that securing AI requires access to both closed and open models, stressing that defenders need the tools to counter emerging threats.The announcement also follows reports that the Trump administration considered restricting access to cutting-edge Chinese models and an industry movement — again spearheaded by Nvidia — defending the need for openness in AI. Google and OpenAI signed that letter belatedly, too, though Anthropic remains absent.#Nvidia #Microsoft #launch #open #security #alliance #OpenAI #Google #AnthropicAI,News,Nvidia,Tech

said it is joining forces with Microsoft, SpaceX, IBM, and other tech companies to build and share open-source AI security tools.

The new Open Secure AI Alliance said open tools are required to effectively defend against attacks from frontier models. The initiative is a direct response to mounting concerns over the safety of advanced AI systems after a rogue OpenAI model escaped containment and attacked another company during testing. That company, Hugging Face, said it was forced to use a Chinese open-weight model to defend itself due to the strict safety guardrails limiting the usefulness of top US models.

Founding members include Palantir, OpenClaw, the Linux Foundation, Cloudflare, Cloudera, Dell, Cisco, Adobe, Siemens, and DoorDash. Conspicuously absent are leading US AI companies, including OpenAI, Google, and Anthropic.

The alliance arrives amid growing tensions over whether the world’s most capable AI models should remain open. Chinese companies have released increasingly powerful open-weight models, notably Moonshot AI’s Kimi K3, challenging the strategy pursued by US labs that have largely kept frontier systems closed and proprietary. Nvidia and its partners argue that securing AI requires access to both closed and open models, stressing that defenders need the tools to counter emerging threats.

The announcement also follows reports that the Trump administration considered restricting access to cutting-edge Chinese models and an industry movement — again spearheaded by Nvidia — defending the need for openness in AI. Google and OpenAI signed that letter belatedly, too, though Anthropic remains absent.

#Nvidia #Microsoft #launch #open #security #alliance #OpenAI #Google #AnthropicAI,News,Nvidia,Tech">Nvidia, Microsoft launch open AI security alliance — without OpenAI, Google, or Anthropic

Nvidia on Monday said it is joining forces with Microsoft, SpaceX, IBM, and other tech companies to build and share open-source AI security tools.

The new Open Secure AI Alliance said open tools are required to effectively defend against attacks from frontier models. The initiative is a direct response to mounting concerns over the safety of advanced AI systems after a rogue OpenAI model escaped containment and attacked another company during testing. That company, Hugging Face, said it was forced to use a Chinese open-weight model to defend itself due to the strict safety guardrails limiting the usefulness of top US models.

Founding members include Palantir, OpenClaw, the Linux Foundation, Cloudflare, Cloudera, Dell, Cisco, Adobe, Siemens, and DoorDash. Conspicuously absent are leading US AI companies, including OpenAI, Google, and Anthropic.

The alliance arrives amid growing tensions over whether the world’s most capable AI models should remain open. Chinese companies have released increasingly powerful open-weight models, notably Moonshot AI’s Kimi K3, challenging the strategy pursued by US labs that have largely kept frontier systems closed and proprietary. Nvidia and its partners argue that securing AI requires access to both closed and open models, stressing that defenders need the tools to counter emerging threats.

The announcement also follows reports that the Trump administration considered restricting access to cutting-edge Chinese models and an industry movement — again spearheaded by Nvidia — defending the need for openness in AI. Google and OpenAI signed that letter belatedly, too, though Anthropic remains absent.

#Nvidia #Microsoft #launch #open #security #alliance #OpenAI #Google #AnthropicAI,News,Nvidia,Tech

Post Comment