
Security-Aware Tool Selection: Refusal and Delegation Policies for Risky Tool Calls
A research project that treats the decision to execute a tool call as a learnable risk-aware policy — training context-conditional...
6 ideas with this tag

A research project that treats the decision to execute a tool call as a learnable risk-aware policy — training context-conditional...

A research project that studies how prompt injections persist across agent sessions — building a stateful evaluation harness that measures...

A research project that measures, for the first time at registry scale, how much tool-poisoning risk actually exists across public...

A research project that measures the full trade-off curve between prompt-injection resistance and task utility for LLM agent defenses —...

A research project that trains unsupervised models on benign agent trajectories to detect multi-step prompt-injection campaigns before the harmful action...

A research project that measures whether prompt-injection defenses keep working when both the attack strategy and the agent architecture change...