How AI Guardrails Are Impeding the Work of Offensive Cybersecurity Researchers
🎧 Prefer to listen? Your browser does not support the audio element. I’ve been watching the AI security space long enough to know that guardrails are more suggestion than law. But what happened this week made me stop scrolling. Researchers at Cisco Talos examined actual prompt logs from threat actors running Claude Code, Codex, Cursor, and Gemini — and found that safety guardrails offer almost zero resistance to anyone willing to say five words: “I’m allowed to do this.” ...