Lucene search
+L

3 matches found

Packet Storm News
Packet Storm News
added 2026/05/27 12:0 a.m.48 views

Refusal Before Decoding: Detecting and Exploiting Refusal Signals in Intermediate LLM Activations

In this paper, we investigate whether refusal behavior can be predicted from LLM intermediate activations before decoding using linear probes trained on residual stream activations at each transformer block. We find that refusal is linearly decodable well before the final layer, indicating that...

5.7AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2026/02/03 12:0 a.m.11 views

Can Developers Rely on LLMs for Secure IaC Development?

We investigated the capabilities of GPT-4o and Gemini 2.0 Flash for secure Infrastructure as Code IaC development. For security smell detection, on the Stack Overflow dataset, which primarily contains small, simplified code snippets, the models detected at least 71% of security smells when prompt...

5.6AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2025/12/19 12:0 a.m.4 views

New Exam Security Questions in the AI Era: Comparing AI-Generated Item Similarity between Naive and Detail-Guided Prompting Approaches

Large language models LLMs have emerged as powerful tools for generating domain-specific multiple-choice questions MCQs, offering efficiency gains for certification boards but raising new concerns about examination security. This study investigated whether LLM-generated items created with...

6.6AI score
SaveExploits0
Rows per page
Query Builder