Lucene search
+L

70 matches found

Packet Storm News
Packet Storm News
added 2025/05/21 12:0 a.m.5 views

Are Vision-Language Models Safe in the Wild? A Meme-Based Benchmark Study

Rapid deployment of vision-language models VLMs magnifies safety risks, yet most evaluations rely on artificial images. This study asks: How safe are current VLMs when confronted with meme images that ordinary users share? To investigate this question, we introduce MemeSafetyBench, a...

7.1AI score
Exploits0
Packet Storm News
Packet Storm News
added 2025/05/21 12:0 a.m.36 views

FragFake: a Dataset for Fine-Grained Detection of Edited Images with Vision Language Models

Fine-grained edited image detection of localized edits in images is crucial for assessing content authenticity, especially given that modern diffusion models and image editing methods can produce highly realistic manipulations. However, this domain faces three challenges: 1 Binary classifiers yie...

6.9AI score
Exploits0
Packet Storm News
Packet Storm News
added 2025/05/21 12:0 a.m.4 views

Blind Spot Navigation: Evolutionary Discovery of Sensitive Semantic Concepts for LVLMs

Whitepaper called Blind Spot Navigation: Evolutionary Discovery Of Sensitive Semantic Concepts For LVLMs...

7AI score
Exploits0
Packet Storm News
Packet Storm News
added 2025/05/20 12:0 a.m.16 views

Beyond Text: Unveiling Privacy Vulnerabilities in Multi-Modal Retrieval-Augmented Generation

Multimodal Retrieval-Augmented Generation MRAG systems enhance LMMs by integrating external multimodal databases, but introduce unexplored privacy vulnerabilities. While text-based RAG privacy risks have been studied, multimodal data presents unique challenges. We provide the first systematic...

6.9AI score
Exploits0
Packet Storm News
Packet Storm News
added 2025/05/16 12:0 a.m.12 views

GuardReasoner-VL: Safeguarding VLMs Via Reinforced Reasoning

To enhance the safety of VLMs, this paper introduces a novel reasoning-based VLM guard model dubbed GuardReasoner-VL. The core idea is to incentivize the guard model to deliberatively reason before making moderation decisions via online RL. First, we construct GuardReasoner-VLTrain, a reasoning...

7.3AI score
Exploits0
Packet Storm News
Packet Storm News
added 2025/05/16 12:0 a.m.6 views

Diverging Towards Hallucination: Detection of Failures in Vision-Language Models Via Multi-Token Aggregation

Vision-language models VLMs now rival human performance on many multimodal tasks, yet they still hallucinate objects or generate unsafe text. Current hallucination detectors, e.g., single-token linear probing SLP and PTrue, typically analyze only the logit of the first generated token or just its...

6.9AI score
Exploits0
Packet Storm News
Packet Storm News
added 2025/05/13 12:0 a.m.11 views

Removing Watermarks with Partial Regeneration Using Semantic Information

As AI-generated imagery becomes ubiquitous, invisible watermarks have emerged as a primary line of defense for copyright and provenance. The newest watermarking schemes embed semantic signals - content-aware patterns that are designed to survive common image manipulations - yet their true...

6.9AI score
Exploits0
Packet Storm News
Packet Storm News
added 2025/05/08 12:0 a.m.5 views

Building Trustworthy Multimodal AI: a Review of Fairness, Transparency, and Ethics in Vision-Language Tasks

Objective: This review explores the trustworthiness of multimodal artificial intelligence AI systems, specifically focusing on vision-language tasks. It addresses critical challenges related to fairness, transparency, and ethical implications in these systems, providing a comparative analysis of...

7AI score
Exploits0
Packet Storm News
Packet Storm News
added 2025/04/25 12:0 a.m.7 views

Revisiting Data Auditing in Large Vision-Language Models

With the surge of large language models LLMs, Large Vision-Language Models VLMs--which integrate vision encoders with LLMs for accurate visual grounding--have shown great potential in tasks like generalist agents and robotic control. However, VLMs are typically trained on massive web-scraped...

6.9AI score
Exploits0
Packet Storm News
Packet Storm News
added 2025/04/15 12:0 a.m.5 views

R-TPT: Improving Adversarial Robustness of Vision-Language Models through Test-Time Prompt Tuning

Vision-language models VLMs, such as CLIP, have gained significant popularity as foundation models, with numerous fine-tuning methods developed to enhance performance on downstream tasks. However, due to their inherent vulnerability and the common practice of selecting from a limited set of...

6.8AI score
Exploits0
Rows per page
Query Builder