1 matches found
deleting-the-trace
Control-token injection attacks on tool-using agents Code and logged measurements for a study of two input-level attacks on tool-using language model agents. The first attack appends a short string of a model's own channel-control tokens to untrusted input; the tokenizer reads it as an...
6AI score
SaveExploits0
20