Lucene search
+L

1 matches found

Packet Storm News
Packet Storm News
added 2025/05/16 12:0 a.m.14 views

GuardReasoner-VL: Safeguarding VLMs Via Reinforced Reasoning

To enhance the safety of VLMs, this paper introduces a novel reasoning-based VLM guard model dubbed GuardReasoner-VL. The core idea is to incentivize the guard model to deliberatively reason before making moderation decisions via online RL. First, we construct GuardReasoner-VLTrain, a reasoning...

7.3AI score
SaveExploits0
Rows per page
Query Builder