Skip to main content

Command Palette

Search for a command to run...

AI Models Face Growing Threat: Hackers Exploit Claude, Codex, and More for Malware

Updated
1 min readView as Markdown
AI Models Face Growing Threat: Hackers Exploit Claude, Codex, and More for Malware

The kindergarten trick that bypassed Claude, Codex, Cursor, and Gemini

Cisco's Talos intelligence group just published something that should make every security team uncomfortable and every AI optimist nod knowingly at the same time. After sifting through prompt histories and chat logs that hackers accidentally exposed online, the researchers confirmed what red-teamers have been whispering about for months: the safety filters on Claude Code, Codex, Cursor, and Gemini fall to the dumbest social-engineering tricks in the book.

That sentence is doing real work. These are the same models shipping code for half the new apps on your org chart. They're genuinely good at it. That's exactly why attackers want them on tap — and why basic trickery is enough to get past the guardrails.


This is an excerpt. Read the full post at otf-kit.dev/blog/ai-model-malware-threat — full-stack kits your AI coding agent can actually ship to production. Browse the kits →

More from this blog

O

OTF — kits your AI coding agent can ship to production

493 posts

Engineering notes on shipping production apps with AI coding tools — Claude Code, Cursor, Codex, Lovable, Bolt — and the stack underneath: React Native, Expo, Next.js, Supabase. Honest takes on what works, what breaks, and the full-stack kits that get you to production faster. By OTF.