Lucene search
+L

1 matches found

Packet Storm News
Packet Storm News
added 2026/05/11 12:0 a.m.20 views

Can You Keep a Secret? Involuntary Information Leakage in Language Model Writing

Language models are deployed in settings that require compartmentalization: system prompts should not be disclosed, chain-of-thought reasoning is hidden from users, and sensitive data passes through shared contexts. We test whether models can keep prompted information out of their writing. We giv...

5.7AI score
SaveExploits0
Rows per page
Query Builder