Lucene search
+L

6 matches found

Kitploit
Kitploit
•added 2026/10/06 11:27 a.m.•10 views

IPI-exposure-signal

IPI Exposure Signal Este es el repositorio de código de nuestro artículo: Your Agentic LLMs Secretly Encode Latent Signals of Indirect Prompt-Injection Exposure. Versión en arXiv y enlace al artículo: https://arxiv.org/abs/2608.02657 Este repositorio implementa el pipeline de sondeo para señales...

6.2AI score
SaveExploits0
Kitploit
Kitploit
•added 2026/10/06 3:59 a.m.•1 views

benign-instruction-bench

Aprobar el examen con el que se entrenó Reevaluando los detectores de inyección de prompts donde los agentes LLM realmente los utilizan: en las salidas de herramientas que un agente lee. Los equipos eligen detectores de inyección por sus puntuaciones en los benchmarks. Comprobamos si esas...

SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/10/03 12:00 a.m.•4 views

The Same Zero: Why Identical ASR Can Imply Different Guarantees in LLM-Agent Security

LLM-agent security has produced a dense landscape of defenses - prompt hardening, content filters, permission gates, sandboxes - yet no framework tells a deployer what a defense actually guarantees, or where that guarantee comes from. We apply Verification Autonomy Levels VAL - L0: LLM...

SaveExploits0
Kitploit
Kitploit
•added 2026/09/30 3:46 p.m.•9 views

ActGuard

ActGuard ActGuard es una defensa de auditoría de acciones previa a la ejecución contra la inyección indirecta de prompts en agentes LLM que utilizan herramientas. Este repositorio contiene la implementación final de ActGuard y el entorno de ejecución basado en AgentDojo necesario para evaluarla. ...

6.3AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/02/11 12:00 a.m.•24 views

Optimizing Agent Planning for Security and Autonomy

Indirect prompt injection attacks threaten AI agents that execute consequential actions, motivating deterministic system-level defenses. Such defenses can provably block unsafe actions by enforcing confidentiality and integrity policies, but currently appear costly: they reduce task completion...

5.6AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2025/04/15 12:00 a.m.•14 views

Progent: Programmable Privilege Control for LLM Agents

LLM agents are an emerging form of AI systems where large language models LLMs serve as the central component, utilizing a diverse set of tools to complete user-assigned tasks. Despite their great potential, LLM agents pose significant security risks. When interacting with the external world, the...

7.3AI score
SaveExploits0
Rows per page
Query Builder