1 matches found
BadEngram: Backdoor Attack on Gated Memory Components in LLMs
To expand open-weight models' capacity without proportionally increasing computation, recent language models incorporate gated parametric memories that retrieve learned values and inject them into intermediate representations. Despite these efficiency benefits, such modules create a distinct atta...
5.9AI score
SaveExploits0
20