Lucene search
+L

3 matches found

Packet Storm News
Packet Storm News
•added 2026/08/20 12:00 a.m.•14 views

AiXamine: Unified Black-Box Evaluation of Cross-Dimensional Trade-Offs in LLM Safety, Security, and Privacy

The critical failure modes in deployed large language models LLMs are cross-dimensional: a model can score 99.3 in safety alignment while refusing one in three benign queries, or improve across every capability metric while losing 21 points in privacy. Existing evaluation frameworks that assess...

5.6AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/08/08 12:00 a.m.•28 views

BASIS: Breach-Aware Selective Prompt Injection Shielding with Prefill Attention Probes

Prompt injection is a critical security threat in large language model LLM applications, where attackers hijack model behavior by embedding malicious instructions in user or external data. Existing detection methods only detect the presence of injection and refuse to respond upon detection,...

5.6AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/07/27 12:00 a.m.•18 views

When LLM Defenses Backfire: Characterizing Safety, Performance, and Cost Trade-Offs

Jailbreak defenses are essential for protecting large language models LLMs, but they can also introduce secondary costs that weaken model utility. We present a systematic study of these defense trade-offs along three dimensions: performance impact, over-refusal on benign inputs, and inference cos...

5.9AI score
SaveExploits0
Rows per page
Query Builder