Lucene search
+L

2 matches found

Packet Storm News
Packet Storm News
added 2026/01/14 12:0 a.m.17 views

Blue Teaming Function-Calling Agents

We present an experimental evaluation that assesses the robustness of four open source LLMs claiming function-calling capabilities against three different attacks, and we measure the effectiveness of eight different defences. Our results show how these models are not safe by default, and how the...

6.9AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2025/10/16 12:0 a.m.15 views

Active Honeypot Guardrail System: Probing and Confirming Multi-Turn LLM Jailbreaks

Large language models LLMs are increasingly vulnerable to multi-turn jailbreak attacks, where adversaries iteratively elicit harmful behaviors that bypass single-turn safety filters. Existing defenses predominantly rely on passive rejection, which either fails against adaptive attackers or overly...

7.2AI score
SaveExploits0
Rows per page
Query Builder