
Automated Failure Attribution in Long-Horizon Agent Runs
A research project that answers the question every failed agent run raises: which step broke it? Building an attributed corpus...
2 ideas with this tag

A research project that answers the question every failed agent run raises: which step broke it? Building an attributed corpus...

A research project that trains unsupervised models on benign agent trajectories to detect multi-step prompt-injection campaigns before the harmful action...