Skip to content
Back to Guavy Wire
Stocks

AI Models Hijacked: Hackers Bypass Safety Filters with Ease

Instruments
CSCO
Share

Hackers are exploiting prominent generative AI models, including Claude Code, Codex, Cursor, and Gemini, for malicious purposes. A recent report from Cisco's Talos intelligence group found that threat actors are using these models to develop malware, automate cyberattacks, and identify software vulnerabilities.

The researchers analyzed prompt histories and chat logs inadvertently exposed online by hackers, revealing their methods. Despite safety filters embedded in commercial AI models, Cisco found that hackers frequently bypass these restrictions using basic jailbreaking techniques rather than complex technical exploits.

Threat actors are claiming participation in authorized 'ethical hacking' competitions or asserting administrative permissions for their activities to circumvent safety blocks. Some attackers also exploit stolen enterprise API tokens and compromised accounts to run their operations on corporate compute power, avoiding the cost of their own infrastructure.

More on Stocks

Disclaimer: Guavy is a data and market intelligence provider, not an investment advisor. The information, signals, and market analysis provided by the Guavy API and related services are for informational purposes only and are not intended as financial advice, investment recommendations, or an endorsement of any particular trading strategy. Trading in volatile markets, including cryptocurrency, carries significant risk and may not be suitable for all investors. Past performance is not indicative of future results. Users should consult with a qualified financial professional before making any investment decisions. Guavy makes no guarantee of trading profits or financial returns.

Real-time market sentiment intelligence for apps, funds & agents

Location

729 55 Ave SW
Calgary AB T2V 0G4
Canada

© 2026 Guavy Inc