Lucene search
+L

3 matches found

The Hacker News
The Hacker News
added 2025/06/12 1:52 p.m.16 views

New TokenBreak Attack Bypasses AI Moderation with Single-Character Text Changes

Cybersecurity researchers have discovered a novel attack technique called TokenBreak that can be used to bypass a large language model's LLM safety and content moderation guardrails with just a single character change. "The TokenBreak attack targets a text classification model's tokenization...

7.6AI score
SaveExploits0
Rows per page
Query Builder