×
X Is Drowning in Disinformation Following US and Israel’s Attack on Iran

X Is Drowning in Disinformation Following US and Israel’s Attack on Iran

Minutes after Donald Trump announced that the US and Israeli governments had launched a “major combat operation” against Iran in the early hours of Saturday morning, disinformation about the attack and Tehran’s response flooded X.

WIRED has reviewed hundreds of posts on X, some of which have racked up millions of views, that promote misleading claims about the locations and scale of the attack.

Elon Musk’s social media platform is a verifiable mess: In some cases, alleged video footage of the attack shared in posts on X are actually months or years old. In several posts, video footage of apparent attacks have been attributed to incorrect locations. A number of images shared on X appear to be altered or generated with AI. Other posts attempt to pass off video game footage as scenes from the conflict.

X did not respond to a request for comment. Under Musk’s stewardship, X has become a haven for disinformation, especially during major global breaking news events. At the beginning of the Israel-Hamas war, and more recently during anti-immigration enforcement protests in LA, the platform has drowned in inaccurate and faulty posts.

Almost all of the most viral posts reviewed by WIRED on Saturday came from accounts with blue check marks, meaning they pay X for its premium service and could be eligible to earn money based on how much engagement their posts generate, even if the content is false. While some posts with disinformation have a community note appended beneath them to correct the record, they remain up on the site, and it’s unclear how many people viewed them before the notes appeared.

One video posted by a blue check mark account claimed to show ballistic missiles over Dubai; the clip actually showed Iranian ballistic missiles fired at Tel Aviv in October 2024. The post has been viewed over 4.4 million times.

One of the most viral clips shared on X in the hours after the attack claims to show an Israeli fighter jet being shot down by Iranian air defense systems. The video has been shared by dozens of accounts, including one post which has been viewed more than 3.5 million times. The provenance of the video is unclear, but there have been no credible reports of any Israeli jets being shot down over Iran on Saturday.

Another account that claims to be an expert in open source intelligence posted a video showing explosions, alongside the caption: “6 Iranian Hypersonic Missiles hit the Indian-invested Israeli Haifa port. Massive damages reported.” The video has been viewed 64,000 times, but the footage was actually captured last July and shows an Israeli attack on the defense ministry in Damascus, Syria.

In a number of cases, pro-Iranian accounts have been using images and footage from Saturday’s attacks to falsely claim successful strikes against Israel. “IRANIAN MISSILE IMPACT IN TEL AVIV RIGHT NOW,” the Iran Observer account wrote in a post featuring an image of Dubai. The post had been viewed over 200,000 times before it was deleted, but dozens of other posts sharing the same image and making the same claims remain on X.

Tehran Times, a news outlet aligned with the Iranian government, posted what appears to be an AI-generated image on X which claims to show that “an American radar in Qatar was completely destroyed today in an Iranian drone strike.” The use of AI generated images was flagged on X by Tal Hagin, a senior analyst with open source intelligence company Golden Owl. While there are reports that drone and missile attacks targeted the US Navy’s 5th Fleet headquarters in Bahrain, there are no reports yet of similar successful attacks in Qatar.

A pro-Trump account, which also features a blue check mark, posted images claiming to show the before and after pictures of the palace of Iranian Supreme Leader Ali Khamenei, which was targeted during Saturday’s missile attacks. (In a post on Truth Social, Trump claimed Khamenei was killed in an attack.) While the after picture appears to accurately show the palace after the attack, the before picture shows the Mausoleum of Ruhollah Khomeini, which is located on the other side of Tehran. The post has been viewed 365,000 times.

Source link
#Drowning #Disinformation #Israels #Attack #Iran

For some social media users, Facebook may feel like an app left behind in an online era of the past. However, for others, Facebook has reinvented itself as the modern-day Craigslist with Facebook Marketplace.

Meta seems to have taken notice and is now releasing a standalone app for the Facebook Marketplace power users who keep the platform going: It’s sellers.

The new app, aptly named Seller, launched on Friday and provides Facebook Marketplace sellers with tools to list their wares and manage their current listings. 

Like Facebook Marketplace itself, the Seller app is free for users. Meta monetizes Facebook Marketplace with ads and does not take any percentage of its users’ sales. Users will be able to download the app and login with their Facebook account starting today.

Meta launches Seller, a standalone app for Facebook Marketplace sellers
                                                            For some social media users, Facebook may feel like an app left behind in an online era of the past. However, for others, Facebook has reinvented itself as the modern-day Craigslist with Facebook Marketplace.Meta seems to have taken notice and is now releasing a standalone app for the Facebook Marketplace power users who keep the platform going: It’s sellers.The new app, aptly named Seller, launched on Friday and provides Facebook Marketplace sellers with tools to list their wares and manage their current listings. 
Like Facebook Marketplace itself, the Seller app is free for users. Meta monetizes Facebook Marketplace with ads and does not take any percentage of its users’ sales. Users will be able to download the app and login with their Facebook account starting today.
        
            Mashable Light Speed
        
        
    

    
                    


            
            
            
            Credit: Facebook
        
    
Seller provides a tool to post new listings, a dashboard to manage their existing ones, an in-app inbox to reply to prospective buyers, and analytics for all of their old listings. Seller also introduces a few new tools for sellers as well. According to Facebook, Seller users will be able to scan their products with their smartphone, and AI will fill-out the listing details automatically.In an interview with the New York Times, Facebook head Tom Allison said that the app would be the Marketplace seller equivalent of a video editing tool for content creators. Allison also explained how Meta is experimenting with more standalone products for Facebook’s power users, which include Marketplace Sellers and Facebook Groups administrators.
As social media users move to competing platforms like TikTok and Meta’s own Instagram, Facebook has found new life within specific, once-niche features on the platform. In May, Facebook launched a standalone app called Forum for its Facebook Groups power users. Now, the company is doing the same for Facebook Marketplace sellers.Meta says that one-third of young adults who use Facebook also use Marketplace. And, every month, more than 420 million items are listed on the Marketplace platform. 

                    
                                            
                            
                        
                                    #Meta #launches #Seller #standalone #app #Facebook #Marketplace #sellers

Credit: Facebook

Seller provides a tool to post new listings, a dashboard to manage their existing ones, an in-app inbox to reply to prospective buyers, and analytics for all of their old listings. 

Seller also introduces a few new tools for sellers as well. According to Facebook, Seller users will be able to scan their products with their smartphone, and AI will fill-out the listing details automatically.

In an interview with the New York Times, Facebook head Tom Allison said that the app would be the Marketplace seller equivalent of a video editing tool for content creators. Allison also explained how Meta is experimenting with more standalone products for Facebook’s power users, which include Marketplace Sellers and Facebook Groups administrators.

As social media users move to competing platforms like TikTok and Meta’s own Instagram, Facebook has found new life within specific, once-niche features on the platform. In May, Facebook launched a standalone app called Forum for its Facebook Groups power users. Now, the company is doing the same for Facebook Marketplace sellers.

Meta says that one-third of young adults who use Facebook also use Marketplace. And, every month, more than 420 million items are listed on the Marketplace platform. 

#Meta #launches #Seller #standalone #app #Facebook #Marketplace #sellers">Meta launches Seller, a standalone app for Facebook Marketplace sellers
                                                            For some social media users, Facebook may feel like an app left behind in an online era of the past. However, for others, Facebook has reinvented itself as the modern-day Craigslist with Facebook Marketplace.Meta seems to have taken notice and is now releasing a standalone app for the Facebook Marketplace power users who keep the platform going: It’s sellers.The new app, aptly named Seller, launched on Friday and provides Facebook Marketplace sellers with tools to list their wares and manage their current listings. 
Like Facebook Marketplace itself, the Seller app is free for users. Meta monetizes Facebook Marketplace with ads and does not take any percentage of its users’ sales. Users will be able to download the app and login with their Facebook account starting today.
        
            Mashable Light Speed
        
        
    

    
                    


            
            
            
            Credit: Facebook
        
    
Seller provides a tool to post new listings, a dashboard to manage their existing ones, an in-app inbox to reply to prospective buyers, and analytics for all of their old listings. Seller also introduces a few new tools for sellers as well. According to Facebook, Seller users will be able to scan their products with their smartphone, and AI will fill-out the listing details automatically.In an interview with the New York Times, Facebook head Tom Allison said that the app would be the Marketplace seller equivalent of a video editing tool for content creators. Allison also explained how Meta is experimenting with more standalone products for Facebook’s power users, which include Marketplace Sellers and Facebook Groups administrators.
As social media users move to competing platforms like TikTok and Meta’s own Instagram, Facebook has found new life within specific, once-niche features on the platform. In May, Facebook launched a standalone app called Forum for its Facebook Groups power users. Now, the company is doing the same for Facebook Marketplace sellers.Meta says that one-third of young adults who use Facebook also use Marketplace. And, every month, more than 420 million items are listed on the Marketplace platform. 

                    
                                            
                            
                        
                                    #Meta #launches #Seller #standalone #app #Facebook #Marketplace #sellers

releasing a standalone app for the Facebook Marketplace power users who keep the platform going: It’s sellers.

The new app, aptly named Seller, launched on Friday and provides Facebook Marketplace sellers with tools to list their wares and manage their current listings. 

Like Facebook Marketplace itself, the Seller app is free for users. Meta monetizes Facebook Marketplace with ads and does not take any percentage of its users’ sales. Users will be able to download the app and login with their Facebook account starting today.

Meta launches Seller, a standalone app for Facebook Marketplace sellers
                                                            For some social media users, Facebook may feel like an app left behind in an online era of the past. However, for others, Facebook has reinvented itself as the modern-day Craigslist with Facebook Marketplace.Meta seems to have taken notice and is now releasing a standalone app for the Facebook Marketplace power users who keep the platform going: It’s sellers.The new app, aptly named Seller, launched on Friday and provides Facebook Marketplace sellers with tools to list their wares and manage their current listings. 
Like Facebook Marketplace itself, the Seller app is free for users. Meta monetizes Facebook Marketplace with ads and does not take any percentage of its users’ sales. Users will be able to download the app and login with their Facebook account starting today.
        
            Mashable Light Speed
        
        
    

    
                    


            
            
            
            Credit: Facebook
        
    
Seller provides a tool to post new listings, a dashboard to manage their existing ones, an in-app inbox to reply to prospective buyers, and analytics for all of their old listings. Seller also introduces a few new tools for sellers as well. According to Facebook, Seller users will be able to scan their products with their smartphone, and AI will fill-out the listing details automatically.In an interview with the New York Times, Facebook head Tom Allison said that the app would be the Marketplace seller equivalent of a video editing tool for content creators. Allison also explained how Meta is experimenting with more standalone products for Facebook’s power users, which include Marketplace Sellers and Facebook Groups administrators.
As social media users move to competing platforms like TikTok and Meta’s own Instagram, Facebook has found new life within specific, once-niche features on the platform. In May, Facebook launched a standalone app called Forum for its Facebook Groups power users. Now, the company is doing the same for Facebook Marketplace sellers.Meta says that one-third of young adults who use Facebook also use Marketplace. And, every month, more than 420 million items are listed on the Marketplace platform. 

                    
                                            
                            
                        
                                    #Meta #launches #Seller #standalone #app #Facebook #Marketplace #sellers

Credit: Facebook

Seller provides a tool to post new listings, a dashboard to manage their existing ones, an in-app inbox to reply to prospective buyers, and analytics for all of their old listings. 

Seller also introduces a few new tools for sellers as well. According to Facebook, Seller users will be able to scan their products with their smartphone, and AI will fill-out the listing details automatically.

In an interview with the New York Times, Facebook head Tom Allison said that the app would be the Marketplace seller equivalent of a video editing tool for content creators. Allison also explained how Meta is experimenting with more standalone products for Facebook’s power users, which include Marketplace Sellers and Facebook Groups administrators.

As social media users move to competing platforms like TikTok and Meta’s own Instagram, Facebook has found new life within specific, once-niche features on the platform. In May, Facebook launched a standalone app called Forum for its Facebook Groups power users. Now, the company is doing the same for Facebook Marketplace sellers.

Meta says that one-third of young adults who use Facebook also use Marketplace. And, every month, more than 420 million items are listed on the Marketplace platform. 

#Meta #launches #Seller #standalone #app #Facebook #Marketplace #sellers">Meta launches Seller, a standalone app for Facebook Marketplace sellers

For some social media users, Facebook may feel like an app left behind in an online era of the past. However, for others, Facebook has reinvented itself as the modern-day Craigslist with Facebook Marketplace.

Meta seems to have taken notice and is now releasing a standalone app for the Facebook Marketplace power users who keep the platform going: It’s sellers.

The new app, aptly named Seller, launched on Friday and provides Facebook Marketplace sellers with tools to list their wares and manage their current listings. 

Like Facebook Marketplace itself, the Seller app is free for users. Meta monetizes Facebook Marketplace with ads and does not take any percentage of its users’ sales. Users will be able to download the app and login with their Facebook account starting today.

Meta launches Seller, a standalone app for Facebook Marketplace sellers
                                                            For some social media users, Facebook may feel like an app left behind in an online era of the past. However, for others, Facebook has reinvented itself as the modern-day Craigslist with Facebook Marketplace.Meta seems to have taken notice and is now releasing a standalone app for the Facebook Marketplace power users who keep the platform going: It’s sellers.The new app, aptly named Seller, launched on Friday and provides Facebook Marketplace sellers with tools to list their wares and manage their current listings. 
Like Facebook Marketplace itself, the Seller app is free for users. Meta monetizes Facebook Marketplace with ads and does not take any percentage of its users’ sales. Users will be able to download the app and login with their Facebook account starting today.
        
            Mashable Light Speed
        
        
    

    
                    


            
            
            
            Credit: Facebook
        
    
Seller provides a tool to post new listings, a dashboard to manage their existing ones, an in-app inbox to reply to prospective buyers, and analytics for all of their old listings. Seller also introduces a few new tools for sellers as well. According to Facebook, Seller users will be able to scan their products with their smartphone, and AI will fill-out the listing details automatically.In an interview with the New York Times, Facebook head Tom Allison said that the app would be the Marketplace seller equivalent of a video editing tool for content creators. Allison also explained how Meta is experimenting with more standalone products for Facebook’s power users, which include Marketplace Sellers and Facebook Groups administrators.
As social media users move to competing platforms like TikTok and Meta’s own Instagram, Facebook has found new life within specific, once-niche features on the platform. In May, Facebook launched a standalone app called Forum for its Facebook Groups power users. Now, the company is doing the same for Facebook Marketplace sellers.Meta says that one-third of young adults who use Facebook also use Marketplace. And, every month, more than 420 million items are listed on the Marketplace platform. 

                    
                                            
                            
                        
                                    #Meta #launches #Seller #standalone #app #Facebook #Marketplace #sellers

Credit: Facebook

Seller provides a tool to post new listings, a dashboard to manage their existing ones, an in-app inbox to reply to prospective buyers, and analytics for all of their old listings. 

Seller also introduces a few new tools for sellers as well. According to Facebook, Seller users will be able to scan their products with their smartphone, and AI will fill-out the listing details automatically.

In an interview with the New York Times, Facebook head Tom Allison said that the app would be the Marketplace seller equivalent of a video editing tool for content creators. Allison also explained how Meta is experimenting with more standalone products for Facebook’s power users, which include Marketplace Sellers and Facebook Groups administrators.

As social media users move to competing platforms like TikTok and Meta’s own Instagram, Facebook has found new life within specific, once-niche features on the platform. In May, Facebook launched a standalone app called Forum for its Facebook Groups power users. Now, the company is doing the same for Facebook Marketplace sellers.

Meta says that one-third of young adults who use Facebook also use Marketplace. And, every month, more than 420 million items are listed on the Marketplace platform. 

#Meta #launches #Seller #standalone #app #Facebook #Marketplace #sellers
For months, AI giants have devised special vetted programs and strict guardrails to limit the use of their models by malicious hackers. But these limits are now hindering the work of legitimate network defenders, as well as that of offensive cybersecurity researchers. 

In June, the U.S. government slapped export control restrictions on Anthropic’s much-hyped AI models Mythos and Fable. The move was prompted at least in part by a report that claimed it was possible to bypass the models’ guardrails designed to prevent users from using them to build and execute malicious cyberattacks.

Regardless of whether the incident was really motivated by fears of a jailbreak, the fact is that Anthropic has repeatedly marketed Mythos as some kind of doomsday cybermachine that can only be given to carefully vetted users, and even then with strict guardrails in place. (The export controls on Fable 5 and Mythos 5 have since been lifted. Fable 5 returned to general access on July 1; Mythos 5 has been reintroduced only to vetted U.S. organizations as part of the government’s review process.)

That kind of gatekeeping isn’t unique to Mythos. Both Anthropic, with its other models, and OpenAI offer cybersecurity researchers programs they can apply to get vetted and — if approved — access models with fewer cybersecurity restrictions: OpenAI’s Trusted Access for Cyber program and Anthropic’s Cyber Verification Program

These guardrails have been widely criticized, particularly by researchers whose job is to find unknown vulnerabilities in systems and devise ways to exploit them before criminals do.

During a recent appearance on a cybersecurity podcast, Mark Dowd, a well-known security researcher, said that, “it’s not really comfortable to me that these random large companies are making arbitrary decisions about what is safe in security and what’s not.”

Dowd has spent decades finding and selling “zero-days” — previously unknown software flaws and the exploits that take advantage of them — to Western governments, rather than reporting them to the software makers so they get patched. Governments pay a premium for vulnerabilities precisely because they stay open, which is useful for intelligence operations.

Dowd admitted his work may make him biased, but he isn’t alone. Several people who work in offensive cybersecurity — they proactively probe systems for weaknesses — described to TechCrunch how they use AI tools and deal with their guardrails. 

Chris Anley, the chief scientist at security consulting giant NCC Group, said that asking an AI model to try to exploit a bug is a key step in confirming it’s a real vulnerability worth fixing. But if a guardrail prompts the model to refuse to answer the question outright, the guardrail hurts defenders, he said.

“This is where the whole offensive versus defensive and guardrails part comes in, because ‘fix this code’ as a prompt is both an essential mechanism for defense but also a roadmap for finding critical vulnerabilities in the code base,” said Anley. “So at the same time, the same tool is both an offensive tool and a defensive tool, and the two can’t really be unpicked.”

It’s “like a hammer,” he continued. “You can’t build a house without a hammer. It’s definitely a tool but it’s also irreducibly a weapon as well.”

When he and his colleagues run into such a roadblock, they sometimes fall back on open source AI models that come with no guardrails at all.

Paolo Stagno, the chief technology officer at Crowdfense, a well-known company that develops, acquires, and sells unknown vulnerabilities to government agencies, agreed with Dowd, saying AI companies “essentially treat customers like children who need babysitting” with their vetted programs and guardrails. 

Stagno said he and his colleagues do use frontier models — but only for reverse engineering. They avoid using AI to help find vulnerabilities or build exploits, he said, because feeding that work into a cloud-based model risks leaking sensitive vulnerability data or having it absorbed into future training runs. For that step, he said, they use open source models run locally, as they do not rely on sharing data outside of the model. 

Giuseppe Cali, a security researcher who finds zero-days and develops exploits, said guardrails are not impeding his work. That’s because he doesn’t use AI for offensive work; instead, he uses it for initial reverse engineering, to understand the code he’s analyzing, and to build supporting tools. For that, he said, AI tools can speed up the process and allow him to focus on discovering vulnerabilities. 

“I still want to own the actual bug discovery and weaponization myself and that wouldn’t change if all guardrails were lifted tomorrow,” said Cali. “I am jealous of my bugs, and I like this game too much to let models play it for me.”

One researcher at a smartphone-component manufacturer, who spoke on condition of anonymity because he isn’t authorized to talk to the press, said his employer isn’t part of Anthropic’s CVP program and as a result, its tools are barely useful for finding vulnerabilities because the guardrails are too strict.

“If it catches wind we’re doing anything security related, it just stops and isn’t usable,” the person said. 

Chris Thompson — chief executive of cybersecurity firm RemoteThreat and founder of Offensive AI Con, an offensive security and AI-focused event — said that in his experience using the frontier AI models, the guardrails can be inconsistent and work differently every day. That’s true even inside the looser boundaries of Anthropic’s and OpenAI’s vetted programs. 

“I think the practical impact is you spend a lot of time negotiating with the model instead of working on the core security program,” said Thompson. “Instead of analyzing a vulnerability and reasoning through the exploitability, you’re trying to find why you’re getting inconsistent results or why are models over-sanitizing the output.” 

Consequently, researchers rely on or get pushed toward Chinese open source models like GLM — freely downloadable models that can be run locally with no vetting or usage restrictions — said Thompson.

“You have these responsible researchers that are being pushed away from U.S.-governed systems to foreign-owned systems,” he said. “I think it’s more harmful than good to have these guardrails in place.”

Rather than tightening restrictions further, Thompson called for the AI frontier labs to open up their programs, provide responsible access, and hold those who abuse their tools accountable. Otherwise, he argued, defenders will lose the AI race.

“There’s this big storm coming. There’s this big wave of attacks that are going to happen at speed and scale like never before,” said Thompson. “But the same security consulting firms and legit researchers that are trying to make a difference are being stifled right now.”

When you purchase through links in our articles, we may earn a small commission. This doesn’t affect our editorial independence.

#guardrails #impeding #work #offensive #cybersecurity #researchers #TechCrunchcybersecurity,Zero-days">How AI guardrails are impeding the work of offensive cybersecurity researchers | TechCrunch
For months, AI giants have devised special vetted programs and strict guardrails to limit the use of their models by malicious hackers. But these limits are now hindering the work of legitimate network defenders, as well as that of offensive cybersecurity researchers. 

In June, the U.S. government slapped export control restrictions on Anthropic’s much-hyped AI models Mythos and Fable. The move was prompted at least in part by a report that claimed it was possible to bypass the models’ guardrails designed to prevent users from using them to build and execute malicious cyberattacks.







Regardless of whether the incident was really motivated by fears of a jailbreak, the fact is that Anthropic has repeatedly marketed Mythos as some kind of doomsday cybermachine that can only be given to carefully vetted users, and even then with strict guardrails in place. (The export controls on Fable 5 and Mythos 5 have since been lifted. Fable 5 returned to general access on July 1; Mythos 5 has been reintroduced only to vetted U.S. organizations as part of the government’s review process.)

That kind of gatekeeping isn’t unique to Mythos. Both Anthropic, with its other models, and OpenAI offer cybersecurity researchers programs they can apply to get vetted and — if approved — access models with fewer cybersecurity restrictions: OpenAI’s Trusted Access for Cyber program and Anthropic’s Cyber Verification Program. 

These guardrails have been widely criticized, particularly by researchers whose job is to find unknown vulnerabilities in systems and devise ways to exploit them before criminals do.

During a recent appearance on a cybersecurity podcast, Mark Dowd, a well-known security researcher, said that, “it’s not really comfortable to me that these random large companies are making arbitrary decisions about what is safe in security and what’s not.”

Dowd has spent decades finding and selling “zero-days” — previously unknown software flaws and the exploits that take advantage of them — to Western governments, rather than reporting them to the software makers so they get patched. Governments pay a premium for vulnerabilities precisely because they stay open, which is useful for intelligence operations.


Dowd admitted his work may make him biased, but he isn’t alone. Several people who work in offensive cybersecurity — they proactively probe systems for weaknesses  — described to TechCrunch how they use AI tools and deal with their guardrails. 

Chris Anley, the chief scientist at security consulting giant NCC Group, said that asking an AI model to try to exploit a bug is a key step in confirming it’s a real vulnerability worth fixing. But if a guardrail prompts the model to refuse to answer the question outright, the guardrail hurts defenders, he said. 

“This is where the whole offensive versus defensive and guardrails part comes in, because ‘fix this code’ as a prompt is both an essential mechanism for defense but also a roadmap for finding critical vulnerabilities in the code base,” said Anley. “So at the same time, the same tool is both an offensive tool and a defensive tool, and the two can’t really be unpicked.” 







It’s “like a hammer,” he continued. “You can’t build a house without a hammer. It’s definitely a tool but it’s also irreducibly a weapon as well.”

When he and his colleagues run into such a roadblock, they sometimes fall back on open source AI models that come with no guardrails at all.

Paolo Stagno, the chief technology officer at Crowdfense, a well-known company that develops, acquires, and sells unknown vulnerabilities to government agencies, agreed with Dowd, saying AI companies “essentially treat customers like children who need babysitting” with their vetted programs and guardrails. 

Stagno said he and his colleagues do use frontier models — but only for reverse engineering. They avoid using AI to help find vulnerabilities or build exploits, he said, because feeding that work into a cloud-based model risks leaking sensitive vulnerability data or having it absorbed into future training runs. For that step, he said, they use open source models run locally, as they do not rely on sharing data outside of the model. 

Giuseppe Cali, a security researcher who finds zero-days and develops exploits, said guardrails are not impeding his work. That’s because he doesn’t use AI for offensive work; instead, he uses it for initial reverse engineering, to understand the code he’s analyzing, and to build supporting tools. For that, he said, AI tools can speed up the process and allow him to focus on discovering vulnerabilities. 

“I still want to own the actual bug discovery and weaponization myself and that wouldn’t change if all guardrails were lifted tomorrow,” said Cali. “I am jealous of my bugs, and I like this game too much to let models play it for me.”

One researcher at a smartphone-component manufacturer, who spoke on condition of anonymity because he isn’t authorized to talk to the press, said his employer isn’t part of Anthropic’s CVP program and as a result, its tools are barely useful for finding vulnerabilities because the guardrails are too strict.

“If it catches wind we’re doing anything security related, it just stops and isn’t usable,” the person said. 







Chris Thompson — chief executive of cybersecurity firm RemoteThreat and founder of Offensive AI Con, an offensive security and AI-focused event — said that in his experience using the frontier AI models, the guardrails can be inconsistent and work differently every day. That’s true even inside the looser boundaries of Anthropic’s and OpenAI’s vetted programs. 

“I think the practical impact is you spend a lot of time negotiating with the model instead of working on the core security program,” said Thompson. “Instead of analyzing a vulnerability and reasoning through the exploitability, you’re trying to find why you’re getting inconsistent results or why are models over-sanitizing the output.” 

Consequently, researchers rely on or get pushed toward Chinese open source models like GLM — freely downloadable models that can be run locally with no vetting or usage restrictions — said Thompson.

“You have these responsible researchers that are being pushed away from U.S.-governed systems to foreign-owned systems,” he said. “I think it’s more harmful than good to have these guardrails in place.”

Rather than tightening restrictions further, Thompson called for the AI frontier labs to open up their programs, provide responsible access, and hold those who abuse their tools accountable. Otherwise, he argued, defenders will lose the AI race.

“There’s this big storm coming. There’s this big wave of attacks that are going to happen at speed and scale like never before,” said Thompson. “But the same security consulting firms and legit researchers that are trying to make a difference are being stifled right now.”


When you purchase through links in our articles, we may earn a small commission. This doesn’t affect our editorial independence.#guardrails #impeding #work #offensive #cybersecurity #researchers #TechCrunchcybersecurity,Zero-days

slapped export control restrictions on Anthropic’s much-hyped AI models Mythos and Fable. The move was prompted at least in part by a report that claimed it was possible to bypass the models’ guardrails designed to prevent users from using them to build and execute malicious cyberattacks.

Regardless of whether the incident was really motivated by fears of a jailbreak, the fact is that Anthropic has repeatedly marketed Mythos as some kind of doomsday cybermachine that can only be given to carefully vetted users, and even then with strict guardrails in place. (The export controls on Fable 5 and Mythos 5 have since been lifted. Fable 5 returned to general access on July 1; Mythos 5 has been reintroduced only to vetted U.S. organizations as part of the government’s review process.)

That kind of gatekeeping isn’t unique to Mythos. Both Anthropic, with its other models, and OpenAI offer cybersecurity researchers programs they can apply to get vetted and — if approved — access models with fewer cybersecurity restrictions: OpenAI’s Trusted Access for Cyber program and Anthropic’s Cyber Verification Program

These guardrails have been widely criticized, particularly by researchers whose job is to find unknown vulnerabilities in systems and devise ways to exploit them before criminals do.

During a recent appearance on a cybersecurity podcast, Mark Dowd, a well-known security researcher, said that, “it’s not really comfortable to me that these random large companies are making arbitrary decisions about what is safe in security and what’s not.”

Dowd has spent decades finding and selling “zero-days” — previously unknown software flaws and the exploits that take advantage of them — to Western governments, rather than reporting them to the software makers so they get patched. Governments pay a premium for vulnerabilities precisely because they stay open, which is useful for intelligence operations.

Dowd admitted his work may make him biased, but he isn’t alone. Several people who work in offensive cybersecurity — they proactively probe systems for weaknesses — described to TechCrunch how they use AI tools and deal with their guardrails. 

Chris Anley, the chief scientist at security consulting giant NCC Group, said that asking an AI model to try to exploit a bug is a key step in confirming it’s a real vulnerability worth fixing. But if a guardrail prompts the model to refuse to answer the question outright, the guardrail hurts defenders, he said.

“This is where the whole offensive versus defensive and guardrails part comes in, because ‘fix this code’ as a prompt is both an essential mechanism for defense but also a roadmap for finding critical vulnerabilities in the code base,” said Anley. “So at the same time, the same tool is both an offensive tool and a defensive tool, and the two can’t really be unpicked.”

It’s “like a hammer,” he continued. “You can’t build a house without a hammer. It’s definitely a tool but it’s also irreducibly a weapon as well.”

When he and his colleagues run into such a roadblock, they sometimes fall back on open source AI models that come with no guardrails at all.

Paolo Stagno, the chief technology officer at Crowdfense, a well-known company that develops, acquires, and sells unknown vulnerabilities to government agencies, agreed with Dowd, saying AI companies “essentially treat customers like children who need babysitting” with their vetted programs and guardrails. 

Stagno said he and his colleagues do use frontier models — but only for reverse engineering. They avoid using AI to help find vulnerabilities or build exploits, he said, because feeding that work into a cloud-based model risks leaking sensitive vulnerability data or having it absorbed into future training runs. For that step, he said, they use open source models run locally, as they do not rely on sharing data outside of the model. 

Giuseppe Cali, a security researcher who finds zero-days and develops exploits, said guardrails are not impeding his work. That’s because he doesn’t use AI for offensive work; instead, he uses it for initial reverse engineering, to understand the code he’s analyzing, and to build supporting tools. For that, he said, AI tools can speed up the process and allow him to focus on discovering vulnerabilities. 

“I still want to own the actual bug discovery and weaponization myself and that wouldn’t change if all guardrails were lifted tomorrow,” said Cali. “I am jealous of my bugs, and I like this game too much to let models play it for me.”

One researcher at a smartphone-component manufacturer, who spoke on condition of anonymity because he isn’t authorized to talk to the press, said his employer isn’t part of Anthropic’s CVP program and as a result, its tools are barely useful for finding vulnerabilities because the guardrails are too strict.

“If it catches wind we’re doing anything security related, it just stops and isn’t usable,” the person said. 

Chris Thompson — chief executive of cybersecurity firm RemoteThreat and founder of Offensive AI Con, an offensive security and AI-focused event — said that in his experience using the frontier AI models, the guardrails can be inconsistent and work differently every day. That’s true even inside the looser boundaries of Anthropic’s and OpenAI’s vetted programs. 

“I think the practical impact is you spend a lot of time negotiating with the model instead of working on the core security program,” said Thompson. “Instead of analyzing a vulnerability and reasoning through the exploitability, you’re trying to find why you’re getting inconsistent results or why are models over-sanitizing the output.” 

Consequently, researchers rely on or get pushed toward Chinese open source models like GLM — freely downloadable models that can be run locally with no vetting or usage restrictions — said Thompson.

“You have these responsible researchers that are being pushed away from U.S.-governed systems to foreign-owned systems,” he said. “I think it’s more harmful than good to have these guardrails in place.”

Rather than tightening restrictions further, Thompson called for the AI frontier labs to open up their programs, provide responsible access, and hold those who abuse their tools accountable. Otherwise, he argued, defenders will lose the AI race.

“There’s this big storm coming. There’s this big wave of attacks that are going to happen at speed and scale like never before,” said Thompson. “But the same security consulting firms and legit researchers that are trying to make a difference are being stifled right now.”

When you purchase through links in our articles, we may earn a small commission. This doesn’t affect our editorial independence.

#guardrails #impeding #work #offensive #cybersecurity #researchers #TechCrunchcybersecurity,Zero-days">How AI guardrails are impeding the work of offensive cybersecurity researchers | TechCrunch

For months, AI giants have devised special vetted programs and strict guardrails to limit the use of their models by malicious hackers. But these limits are now hindering the work of legitimate network defenders, as well as that of offensive cybersecurity researchers. 

In June, the U.S. government slapped export control restrictions on Anthropic’s much-hyped AI models Mythos and Fable. The move was prompted at least in part by a report that claimed it was possible to bypass the models’ guardrails designed to prevent users from using them to build and execute malicious cyberattacks.

Regardless of whether the incident was really motivated by fears of a jailbreak, the fact is that Anthropic has repeatedly marketed Mythos as some kind of doomsday cybermachine that can only be given to carefully vetted users, and even then with strict guardrails in place. (The export controls on Fable 5 and Mythos 5 have since been lifted. Fable 5 returned to general access on July 1; Mythos 5 has been reintroduced only to vetted U.S. organizations as part of the government’s review process.)

That kind of gatekeeping isn’t unique to Mythos. Both Anthropic, with its other models, and OpenAI offer cybersecurity researchers programs they can apply to get vetted and — if approved — access models with fewer cybersecurity restrictions: OpenAI’s Trusted Access for Cyber program and Anthropic’s Cyber Verification Program

These guardrails have been widely criticized, particularly by researchers whose job is to find unknown vulnerabilities in systems and devise ways to exploit them before criminals do.

During a recent appearance on a cybersecurity podcast, Mark Dowd, a well-known security researcher, said that, “it’s not really comfortable to me that these random large companies are making arbitrary decisions about what is safe in security and what’s not.”

Dowd has spent decades finding and selling “zero-days” — previously unknown software flaws and the exploits that take advantage of them — to Western governments, rather than reporting them to the software makers so they get patched. Governments pay a premium for vulnerabilities precisely because they stay open, which is useful for intelligence operations.

Dowd admitted his work may make him biased, but he isn’t alone. Several people who work in offensive cybersecurity — they proactively probe systems for weaknesses — described to TechCrunch how they use AI tools and deal with their guardrails. 

Chris Anley, the chief scientist at security consulting giant NCC Group, said that asking an AI model to try to exploit a bug is a key step in confirming it’s a real vulnerability worth fixing. But if a guardrail prompts the model to refuse to answer the question outright, the guardrail hurts defenders, he said.

“This is where the whole offensive versus defensive and guardrails part comes in, because ‘fix this code’ as a prompt is both an essential mechanism for defense but also a roadmap for finding critical vulnerabilities in the code base,” said Anley. “So at the same time, the same tool is both an offensive tool and a defensive tool, and the two can’t really be unpicked.”

It’s “like a hammer,” he continued. “You can’t build a house without a hammer. It’s definitely a tool but it’s also irreducibly a weapon as well.”

When he and his colleagues run into such a roadblock, they sometimes fall back on open source AI models that come with no guardrails at all.

Paolo Stagno, the chief technology officer at Crowdfense, a well-known company that develops, acquires, and sells unknown vulnerabilities to government agencies, agreed with Dowd, saying AI companies “essentially treat customers like children who need babysitting” with their vetted programs and guardrails. 

Stagno said he and his colleagues do use frontier models — but only for reverse engineering. They avoid using AI to help find vulnerabilities or build exploits, he said, because feeding that work into a cloud-based model risks leaking sensitive vulnerability data or having it absorbed into future training runs. For that step, he said, they use open source models run locally, as they do not rely on sharing data outside of the model. 

Giuseppe Cali, a security researcher who finds zero-days and develops exploits, said guardrails are not impeding his work. That’s because he doesn’t use AI for offensive work; instead, he uses it for initial reverse engineering, to understand the code he’s analyzing, and to build supporting tools. For that, he said, AI tools can speed up the process and allow him to focus on discovering vulnerabilities. 

“I still want to own the actual bug discovery and weaponization myself and that wouldn’t change if all guardrails were lifted tomorrow,” said Cali. “I am jealous of my bugs, and I like this game too much to let models play it for me.”

One researcher at a smartphone-component manufacturer, who spoke on condition of anonymity because he isn’t authorized to talk to the press, said his employer isn’t part of Anthropic’s CVP program and as a result, its tools are barely useful for finding vulnerabilities because the guardrails are too strict.

“If it catches wind we’re doing anything security related, it just stops and isn’t usable,” the person said. 

Chris Thompson — chief executive of cybersecurity firm RemoteThreat and founder of Offensive AI Con, an offensive security and AI-focused event — said that in his experience using the frontier AI models, the guardrails can be inconsistent and work differently every day. That’s true even inside the looser boundaries of Anthropic’s and OpenAI’s vetted programs. 

“I think the practical impact is you spend a lot of time negotiating with the model instead of working on the core security program,” said Thompson. “Instead of analyzing a vulnerability and reasoning through the exploitability, you’re trying to find why you’re getting inconsistent results or why are models over-sanitizing the output.” 

Consequently, researchers rely on or get pushed toward Chinese open source models like GLM — freely downloadable models that can be run locally with no vetting or usage restrictions — said Thompson.

“You have these responsible researchers that are being pushed away from U.S.-governed systems to foreign-owned systems,” he said. “I think it’s more harmful than good to have these guardrails in place.”

Rather than tightening restrictions further, Thompson called for the AI frontier labs to open up their programs, provide responsible access, and hold those who abuse their tools accountable. Otherwise, he argued, defenders will lose the AI race.

“There’s this big storm coming. There’s this big wave of attacks that are going to happen at speed and scale like never before,” said Thompson. “But the same security consulting firms and legit researchers that are trying to make a difference are being stifled right now.”

When you purchase through links in our articles, we may earn a small commission. This doesn’t affect our editorial independence.

#guardrails #impeding #work #offensive #cybersecurity #researchers #TechCrunchcybersecurity,Zero-days

Post Comment