1 matches found
COPEX: Benchmarking LLM Robustness to Adversarial Context across Model Context Protocol Layers
Large language models increasingly mediate tool use in Model Context Protocol MCP systems, where adversarial influence may enter through user instructions, tool schemas, tool outputs, or protocol messages. Existing benchmarks often evaluate deployed agents, conflating model susceptibility with...
5.9AI score
SaveExploits0
20