Lucene search
+L

2 matches found

Packet Storm News
Packet Storm News
added 2026/02/26 12:0 a.m.20 views

Reverse CAPTCHA: Evaluating LLM Susceptibility to Invisible Unicode Instruction Injection

We introduce Reverse CAPTCHA, an evaluation framework that tests whether large language models follow invisible Unicode-encoded instructions embedded in otherwise normal-looking text. Unlike traditional CAPTCHAs that distinguish humans from machines, our benchmark exploits a capability gap: model...

5.8AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2025/06/19 12:0 a.m.8 views

Watermarking Autoregressive Image Generation

Watermarking the outputs of generative models has emerged as a promising approach for tracking their provenance. Despite significant interest in autoregressive image generation models and their potential for misuse, no prior work has attempted to watermark their outputs at the token level. In thi...

6.9AI score
SaveExploits0
Rows per page
Query Builder