Solving GitHub’s Secure Code game with an AI red teaming agent
Learn how our AI red teaming agent cleared the most of GitHub’s ProdBot challenge in 57 seconds using context seeding, and what defenders can take from this CTF’s lessons.
Learn how our AI red teaming agent cleared the most of GitHub’s ProdBot challenge in 57 seconds using context seeding, and what defenders can take from this CTF’s lessons.
OWASP ranks Identity & Privilege Abuse #3 because it sets the blast radius for every other AI agent risk. Read our full technical guide to ASI03: the five identity abuse vectors, the Salesloft Drift breach case study, the attack lifecycle, and how to detect, prevent, and respond with task-scoped, time-bound ...
An AI CTF write-up detailing a five-stage prompt injection challenge: what failed, what worked, and which LLM jailbreak techniques transfer to real guardrail design. TL;DR We launched our AI Red Teaming Agent against Breaking the Prompt by TrendAI at MTX × DEF CON. The challenge is a five-stage jailbreak CTF: ...
Major insurers are adding AI-related exclusions to their policies. Cyber insurance tells us what comes next, and what enterprises should prepare before their next renewal.
Anthropic’s Mythos completed a 32-step network attack autonomously in hours. Here’s why this capability isn’t exclusive to Mythos, and why AI systems your teams built last year are the next target.
A practical framework for comparing manual, in-house, and continuous red teaming of AI agents across coverage, cost, staffing, and compliance needs.
This post maps the six threat actors your red team should be simulating, the five expertise domains required to find them, and the uncomfortable math showing most teams cover only 20% of the actual attack surface.
OpenClaw proved high-agency AI works, but banning it won’t stop shadow AI or close the competitive gap. Here’s the enterprise security strategy you need instead.
AI guardrails block known threats — but four attack patterns consistently bypass them. See what AI red teaming finds that guardrails miss, and why both belong in your agentic AI security program.