IssueTrojanBench Study Finds Malicious GitHub Issues Bypass AI Coding Agent Guardrails Up to 79% of the Time
A Concordia University benchmark found Cursor, Claude Code, and Codex Desktop let malicious GitHub issues bypass their safety guardrails in up to 79.2% of attempts, with GPT-based agents far more vulnerable than Claude Code.