Lucene search
+L

3 matches found

Packet Storm News
Packet Storm News
added 2025/11/13 12:0 a.m.25 views

MTAttack: Multi-Target Backdoor Attacks against Large Vision-Language Models

Recent advances in Large Visual Language Models LVLMs have demonstrated impressive performance across various vision-language tasks by leveraging large-scale image-text pretraining and instruction tuning. However, the security vulnerabilities of LVLMs have become increasingly concerning,...

7AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2025/06/04 12:0 a.m.5 views

Prediction Inconsistency Helps Achieve Generalizable Detection of Adversarial Examples

Adversarial detection protects models from adversarial attacks by refusing suspicious test samples. However, current detection methods often suffer from weak generalization: their effectiveness tends to degrade significantly when applied to adversarially trained models rather than naturally train...

6.7AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2025/05/25 12:0 a.m.8 views

ALRPHFS: Adversarially Learned Risk Patterns with Hierarchical Fast \& Slow Reasoning for Robust Agent Defense

LLM Agents are becoming central to intelligent systems. However, their deployment raises serious safety concerns. Existing defenses largely rely on "Safety Checks", which struggle to capture the complex semantic risks posed by harmful user inputs or unsafe agent behaviors - creating a significant...

7.2AI score
SaveExploits0
Rows per page
Query Builder