Connect with us

News

How AI guardrails are impeding the work of offensive cybersecurity researchers

info

Published

on

Claude mythos logo.jpg

For months, AI giants have devised special vetted programs and strict guardrails to limit the use of their models by malicious hackers. But these limits are now hindering the work of legitimate network defenders, as well as that of offensive cybersecurity researchers. 

In June, the U.S. government slapped export control restrictions on Anthropic’s much-hyped AI models Mythos and Fable. The move was prompted at least in part by a report that claimed it was possible to bypass the models’ guardrails designed to prevent users from using them to build and execute malicious cyberattacks.

Regardless of whether the incident was really motivated by fears of a jailbreak, the fact is that Anthropic has repeatedly marketed Mythos as some kind of doomsday cybermachine that can only be given to carefully vetted users, and even then with strict guardrails in place. (The export controls on Fable 5 and Mythos 5 have since been lifted. Fable 5 returned to general access on July 1; Mythos 5 has been reintroduced only to vetted U.S. organizations as part of the government’s review process.)

That kind of gatekeeping isn’t unique to Mythos. Both Anthropic, with its other models, and OpenAI offer cybersecurity researchers programs they can apply to get vetted and — if approved — access models with fewer cybersecurity restrictions: OpenAI’s Trusted Access for Cyber program and Anthropic’s Cyber Verification Program. 

These guardrails have been widely criticized, particularly by researchers whose job is to find unknown vulnerabilities in systems and devise ways to exploit them before criminals do.

During a recent appearance on a cybersecurity podcast, Mark Dowd, a well-known security researcher, said that, “it’s not really comfortable to me that these random large companies are making arbitrary decisions about what is safe in security and what’s not.”

Dowd has spent decades finding and selling “zero-days” — previously unknown software flaws and the exploits that take advantage of them — to Western governments, rather than reporting them to the software makers so they get patched. Governments pay a premium for vulnerabilities precisely because they stay open, which is useful for intelligence operations.

Dowd admitted his work may make him biased, but he isn’t alone. Several people who work in offensive cybersecurity — they proactively probe systems for weaknesses — described to TechCrunch how they use AI tools and deal with their guardrails. 

Chris Anley, the chief scientist at security consulting giant NCC Group, said that asking an AI model to try to exploit a bug is a key step in confirming it’s a real vulnerability worth fixing. But if a guardrail prompts the model to refuse to answer the question outright, the guardrail hurts defenders, he said.

“This is where the whole offensive versus defensive and guardrails part comes in, because ‘fix this code’ as a prompt is both an essential mechanism for defense but also a roadmap for finding critical vulnerabilities in the code base,” said Anley. “So at the same time, the same tool is both an offensive tool and a defensive tool, and the two can’t really be unpicked.”

It’s “like a hammer,” he continued. “You can’t build a house without a hammer. It’s definitely a tool but it’s also irreducibly a weapon as well.”

When he and his colleagues run into such a roadblock, they sometimes fall back on open source AI models that come with no guardrails at all.

Paolo Stagno, the chief technology officer at Crowdfense, a well-known company that develops, acquires, and sells unknown vulnerabilities to government agencies, agreed with Dowd, saying AI companies “essentially treat customers like children who need babysitting” with their vetted programs and guardrails. 

Stagno said he and his colleagues do use frontier models — but only for reverse engineering. They avoid using AI to help find vulnerabilities or build exploits, he said, because feeding that work into a cloud-based model risks leaking sensitive vulnerability data or having it absorbed into future training runs. For that step, he said, they use open source models run locally, as they do not rely on sharing data outside of the model. 

Giuseppe Cali, a security researcher who finds zero-days and develops exploits, said guardrails are not impeding his work. That’s because he doesn’t use AI for offensive work; instead, he uses it for initial reverse engineering, to understand the code he’s analyzing, and to build supporting tools. For that, he said, AI tools can speed up the process and allow him to focus on discovering vulnerabilities. 

“I still want to own the actual bug discovery and weaponization myself and that wouldn’t change if all guardrails were lifted tomorrow,” said Cali. “I am jealous of my bugs, and I like this game too much to let models play it for me.”

One researcher at a smartphone-component manufacturer, who spoke on condition of anonymity because he isn’t authorized to talk to the press, said his employer isn’t part of Anthropic’s CVP program and as a result, its tools are barely useful for finding vulnerabilities because the guardrails are too strict.

“If it catches wind we’re doing anything security related, it just stops and isn’t usable,” the person said. 

Chris Thompson — chief executive of cybersecurity firm RemoteThreat and founder of Offensive AI Con, an offensive security and AI-focused event — said that in his experience using the frontier AI models, the guardrails can be inconsistent and work differently every day. That’s true even inside the looser boundaries of Anthropic’s and OpenAI’s vetted programs. 

“I think the practical impact is you spend a lot of time negotiating with the model instead of working on the core security program,” said Thompson. “Instead of analyzing a vulnerability and reasoning through the exploitability, you’re trying to find why you’re getting inconsistent results or why are models over-sanitizing the output.” 

Consequently, researchers rely on or get pushed toward Chinese open source models like GLM — freely downloadable models that can be run locally with no vetting or usage restrictions — said Thompson.

“You have these responsible researchers that are being pushed away from U.S.-governed systems to foreign-owned systems,” he said. “I think it’s more harmful than good to have these guardrails in place.”

Rather than tightening restrictions further, Thompson called for the AI frontier labs to open up their programs, provide responsible access, and hold those who abuse their tools accountable. Otherwise, he argued, defenders will lose the AI race.

“There’s this big storm coming. There’s this big wave of attacks that are going to happen at speed and scale like never before,” said Thompson. “But the same security consulting firms and legit researchers that are trying to make a difference are being stifled right now.”

When you purchase through links in our articles, we may earn a small commission. This doesn’t affect our editorial independence.

Continue Reading
Click to comment

Leave a Reply

Your email address will not be published. Required fields are marked *

News

Okpebholo should resign over UK fuel price comparison – APM

info

Published

on

APM.jpg

The Allied Peoples Movement (APM) has called on Edo State Governor, Senator Monday Okpebholo, to resign over his comparison of petrol prices in Nigeria and the United Kingdom.

The party said the governor’s comparison failed to take into account the wide difference in income and purchasing power between Nigerians and people living in the UK.

In a statement by its National Publicity Secretary, Abubakar Yusuf, the APM also criticised the increase in petrol prices to more than N1,500 per litre, saying Nigerians were expecting a reduction in the cost of fuel.

Yusuf said the APM condemned the sharp rise in petrol prices to more than N1,500 per litre, stressing that Nigerians were expecting the price to come down.

Recommended

“It is unfair to compare petrol prices in Nigeria with those in the UK without taking into account the difference in workers’ incomes. Nigeria’s N70,000 minimum wage is far below what workers earn on average in the UK,” the party said.

The APM used teachers to illustrate the difference, saying Nigerian teachers earn about N70,000 or less, while a teacher in the UK starts with a monthly salary of about £2,839.

The party also argued that the two countries have different transport systems. It said Nigerians largely depend on informal and costly transportation, while the UK has a more structured and regulated public transport system supported by government subsidies.

The APM further pointed to other oil-producing countries, including Libya and Iran, where it said petrol is sold at much lower prices.

“What Governor Okpebholo failed to tell Nigerians is the fact that in other oil producing countries such as Libya and even Iran, which is presently at war, petrol is sold at an average of equivalent of N41 to N52 per litre,” the party said.

The party described the governor’s comparison between Nigeria and the UK as provocative, arguing that it could further increase public anger over the rising cost of living.

“It is therefore completely irrational and highly provocative for APC Governor Okpebholo to attempt to compare the situation in the UK to that of Nigeria,” the APM said.

The party also expressed concern that the governor’s comments could be seen as preparing Nigerians for another increase in petrol prices, although it did not provide evidence that such an increase was being planned by the government.

The APM also called on Nigerians to support its presidential candidate, Oyo State Governor Seyi Makinde, in the 2027 election.

Continue Reading

News

Cross River Ikan Sports Fight Back To Beat Gateway Horns 25-19 In Showtime Bowl Series XV Flag Football

info

Published

on

WhatsApp Image 2026 09 27 at 5.22.03 PM.jpeg

Cross River Ikan Sports bounced back from their heavy opening-week defeat with a spirited second-half performance to beat Gateway Horns 25-19 in Week 2 of the Showtime Bowl Series XV Flag Football season at Showtime Arena on Sunday.

After suffering a 56-13 defeat to Lagos Rebels in Week 1, Ikan Sports needed a response and made the perfect start when quarterback Timothy Atibile connected with Praise Miene for a 28-yard touchdown to establish a 6-0 lead.

Read Also: Rivers Alphas Bounce Back, Edge Delta Braves 31-24 In Showtime Bowl Series XV Flag Football Thriller | Sports247 Nigeria

Gateway Horns, however, gradually found their rhythm. Awosika Oluwafikunmi found Adewole Mosimiloluwa for the tying touchdown before connecting with Odimbu Awele Gift for another score. A successful extra point to Mosimiloluwa gave the Horns a 13-6 advantage.

Ikan Sports immediately hit back through one of the biggest plays of the contest as Atibile connected with Abayomi Agbayewa for a 45-yard touchdown, reducing the deficit to 13-12 before halftime.

Gateway entered the break with the narrow advantage, but Ikan Sports turned the contest around in the second half, scoring 13 points while restricting the Horns to six to complete a 25-19 comeback victory and secure their first win of the season.

Praise Miene was named Match MVP after finishing with 33 receiving yards and a touchdown, earning a 7.4 rating. Atibile also played a central role in the victory and topped the fantasy standings with 4.3 points, while Gateway quarterback Oluwafikunmi recorded 4.0.

The result represents an important recovery for Cross River Ikan Sports following their difficult Week 1 outing, while Gateway Horns will be left disappointed after surrendering a 13-12 halftime advantage.

For Ikan Sports, however, Week 2 delivered exactly what they needed — a response, a comeback and their first victory of the Showtime Bowl Series XV campaign.

Continue Reading

Trending