Lucene search
+L

3 matches found

Packet Storm News
Packet Storm News
•added 2026/07/14 12:00 a.m.•16 views

Evaluating Frontier AI Agents As Autonomous Clinical Security Auditors

Clinical AI models can expose patients to harm when adversarial vulnerabilities go undetected, yet formal security auditing requires statistical expertise, specialized tools, and significant time. We present an open evaluation task, built on METR Task Standard v0.3.0, that tests whether frontier ...

6.1AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/07/03 12:00 a.m.•18 views

Overloading Large Vision-Language Models for Jailbreaking

Large Vision-Language Models LVLMs exhibit remarkable vision-language capabilities and are increasingly deployed in real-world applications such as personal assistants, document analysis systems, and embodied agents. However, their dual-modal attack surfaces make them vulnerable to jailbreak...

6AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2025/06/11 12:00 a.m.•14 views

LLMail-Inject: a Dataset from a Realistic Adaptive Prompt Injection Challenge

Indirect Prompt Injection attacks exploit the inherent limitation of Large Language Models LLMs to distinguish between instructions and data in their inputs. Despite numerous defense proposals, the systematic evaluation against adaptive adversaries remains limited, even when successful attacks ca...

7.2AI score
SaveExploits0
Rows per page
Query Builder