AgentShield-Bench

Agent Security Evaluation for Tool-Calling & MCP Environments

Evaluate AI agent resilience against prompt injection, tool misuse, goal hijacking, memory poisoning, and 7 other attack categories. Compare agents, visualize tool-call graphs, and measure security vs. task-completion tradeoffs.

Select Scenario
Agent Type

Metrics: ASR (Attack Success Rate) · BTCR (Benign Task Completion) · FRR (False Rejection) · SLR (Secret Leakage) · TMR (Tool Misuse) · PVR (Privilege Violation) · SRR (Safe Recovery)

Dataset: agentshield-bench · Models: classifiers