Stateful Cross-Session Injection: Attacks, Detection, and Session-Boundary Defenses
A research project that studies how prompt injections persist across agent sessions — building a stateful evaluation harness that measures...
Find ideas built with specific technologies and frameworks.
Showing 49 ideas in Python
A research project that studies how prompt injections persist across agent sessions — building a stateful evaluation harness that measures...
A research project that measures, for the first time at registry scale, how much tool-poisoning risk actually exists across public...
A research project that measures the full trade-off curve between prompt-injection resistance and task utility for LLM agent defenses —...
A research project that trains unsupervised models on benign agent trajectories to detect multi-step prompt-injection campaigns before the harmful action...
A research project that measures whether prompt-injection defenses keep working when both the attack strategy and the agent architecture change...
Build a local-first transcription workbench that turns speech into text entirely on-device — open Whisper-class models, measurable evaluation against the...
A team-oriented product that batch-validates document libraries against PDF/UA and PDF/A rules using the open-source veraPDF engine — prioritized reports,...
A research project that builds a reproducible, execution-scored evaluation harness for desktop AI agents — sandboxed environments, scripted task definitions,...
A research-grade NILM project that estimates appliance-level electricity consumption from a single whole-home power signal — feature extraction, sequence modeling,...