1 matches found
AKRASIA: Stealthy Backdoor Attack on Reasoning-Based Code LLMs
We present AKRASIA, a stealthy, inference-time backdoor attack against reasoning-based Code LLMs. AKRASIA aims to achieve a backdoor target e.g., malicious code execution in reasoning LLMs while evading automated defenses and human inspection. To achieve this, AKRASIA probes the victim LLM to...
5.8AI score
SaveExploits0
20