
Security-Utility Pareto Measurement of Agent Defenses under Adaptive Attacks
A research project that measures the full trade-off curve between prompt-injection resistance and task utility for LLM agent defenses —...
2 ideas with this tag

A research project that measures the full trade-off curve between prompt-injection resistance and task utility for LLM agent defenses —...

A research project that builds a reproducible, execution-scored evaluation harness for desktop AI agents — sandboxed environments, scripted task definitions,...