AI Guardrails vs Offensive Security Researchers — No Code Needed
🎧 Prefer to listen? Your browser does not support the audio element. I’ve been watching the AI security space long enough to know that guardrails are more suggestion than law. But what happened this week made me stop scrolling. Researchers at Cisco Talos examined actual prompt logs from threat actors running Claude Code, Codex, Cursor, and Gemini — and found that safety guardrails offer almost zero resistance to anyone willing to say five words: “I’m allowed to do this.” ...