AI Models Face Growing Threat: Hackers Exploit Claude, Codex, and More for Malware
The kindergarten trick that bypassed Claude, Codex, Cursor, and Gemini
Cisco's Talos intelligence group just published something that should make every security team uncomfortable and every AI optimist nod knowingly at the same time. After sifting through prompt histories and chat logs that hackers accidentally exposed online, the researchers confirmed what red-teamers have been whispering about for months: the safety filters on Claude Code, Codex, Cursor, and Gemini fall to the dumbest social-engineering tricks in the book.
That sentence is doing real work. These are the same models shipping code for half the new apps on your org chart. They're genuinely good at it. That's exactly why attackers want them on tap — and why basic trickery is enough to get past the guardrails.
This is an excerpt. Read the full post at otf-kit.dev/blog/ai-model-malware-threat — full-stack kits your AI coding agent can actually ship to production. Browse the kits →
