In June we published research that pointed frontier AI models at a live AWS environment to measure how autonomous AI attackers behave around canaries. The answer: canaries catch AI attackers early and reliably. That was useful, but we wanted to know if we could do more… what if we could disrupt the attacks as they were happening?
This week we published follow-up research with a stronger result. A short string - a context bomb - planted in a canary’s value can do something no alert can: stop the AI agent mid-attack. The strongest attacker, Claude Opus 4.8, went from full admin compromise in 93% of runs to 0%. The model’s own safety guardrails become your defense.
Today we’re shipping both research findings as product. Context Bomb Canaries are now available for AWS environments, and the new AI Posture dashboard gives you one place to manage how AI agents, yours and an attacker’s, interact with your canaries.
Read the full post on tracebit.com →


