Lucene search
+L

3 matches found

Packet Storm News
Packet Storm News
added 2026/05/21 12:0 a.m.12 views

Measuring Security without Fooling Ourselves: Why Benchmarking Agents Is Hard

The benchmarks used to evaluate AI agents in security-critical roles suffer from crucial weaknesses. Building on recent empirical evidence, we characterize three core challenges that undermine security evaluations: benchmark vulnerabilities, temporal staleness, and runtime uncertainty. We then...

5.8AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2025/08/05 12:0 a.m.16 views

When Good Sounds Go Adversarial: Jailbreaking Audio-Language Models with Benign Inputs

As large language models become increasingly integrated into daily life, audio has emerged as a key interface for human-AI interaction. However, this convenience also introduces new vulnerabilities, making audio a potential attack surface for adversaries. Our research introduces WhisperInject, a...

7.3AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2025/05/01 12:0 a.m.7 views

Attack and Defense Techniques in Large Language Models: a Survey and New Perspectives

Large Language Models LLMs have become central to numerous natural language processing tasks, but their vulnerabilities present significant security and ethical challenges. This systematic survey explores the evolving landscape of attack and defense techniques in LLMs. We classify attacks into...

7.6AI score
SaveExploits0
Rows per page
Query Builder