Problem: MCP guard tools claim to block attacks on AI agents, but there was no consistent way to hold them accountable.
What I built: A benchmark with a sandboxed testing harness that runs guard tools against 6 classes of attacks in a controlled environment, so their protection can be measured and compared fairly.
Skills: AI security, MCP, adversarial testing, Python, sandboxing, benchmarking
Sentinel: MCP Security Guard Benchmark
Problem: MCP guard tools claim to block attacks on AI agents, but there was no consistent way to hold them accountable...