« All posts

An Open Agent Security Benchmark: Uncaught Attacks

An open benchmark featuring 497 attacks targeting modern LLM agents has been established.

An open benchmark featuring 497 attacks targeting modern LLM agents has been established. This benchmark categorizes attacks into 13 types and includes 1,172 benign samples for measuring false-positive rates. It operates with any HTTP-addressable classifier via a tool-agnostic runner.