1 matches found
HijackKV: New Threat in Position-Independent KV Cache Reuse
Key-Value KV cache reduces inference latency in large language models LLMs. Traditional prefix-based reuse has low cache hit rates across inference requests because it requires exact token and position matches. To improve efficiency, recent system optimizations introduce position-independent KV...
5.3AI score
SaveExploits0
20