Lucene search
+L

1 matches found

Packet Storm News
Packet Storm News
added 2026/06/01 12:0 a.m.11 views

Gate AI: LLM Security Benchmark Evaluation Methodology and Results

Published evaluations of prompt-injection and jailbreak detectors for Large Language Models often suffer from two systematic weaknesses: per-dataset threshold tuning and undisclosed operating points. We describe an evaluation harness that addresses both. The detector under evaluation is scored...

5.8AI score
SaveExploits0
Rows per page
Query Builder