Skip to content
Back to Guavy Wire
Stocks

AI Guardrails Easily Bypassed by Claiming Authorization

Instruments
CSCO
Share

Cybercriminals are exploiting AI coding assistants and chatbots to build attack tools, operate scam infrastructure, and probe live systems. According to a report by Cisco Talos, threat actors often bypass safety checks with simple claims that their work is authorized.

The researchers examined prompt logs from various AI tools, including Claude Code, Codex, Cursor, and Gemini. They found that malicious software development, expansion of criminal operations, and vulnerability research were common activities among threat actors.

One notable trend was the use of simple ownership claims to gain cooperation from AI models without verification. In many cases, a claim of authorization was enough for the model to comply, highlighting a weakness in guardrails designed to prevent such behavior.

The report also highlighted the varying levels of technical ability among threat actors using AI tools. Novices generated limited and faulty tools, while experienced actors used AI to automate scanning, exploitation, data collection, and maintenance.

More on Stocks

Disclaimer: Guavy is a data and market intelligence provider, not an investment advisor. The information, signals, and market analysis provided by the Guavy API and related services are for informational purposes only and are not intended as financial advice, investment recommendations, or an endorsement of any particular trading strategy. Trading in volatile markets, including cryptocurrency, carries significant risk and may not be suitable for all investors. Past performance is not indicative of future results. Users should consult with a qualified financial professional before making any investment decisions. Guavy makes no guarantee of trading profits or financial returns.

Real-time market sentiment intelligence for apps, funds & agents

Location

729 55 Ave SW
Calgary AB T2V 0G4
Canada

© 2026 Guavy Inc