5395 matches found
CVE-2025-64495-POC
CVE-2025-64495-POC Open WebUI es vulnerable a DOM XSS almacenado a través de prompts cuando 'Insertar Prompt como Texto Enriquecido' está habilitado, resultando en ATO/RCE Resumen La funcionalidad que inserta prompts personalizados en la ventana de chat es vulnerable a DOM XSS cuando 'Insertar...
sk-cve-2026-26030-lab
CVE-2026-26030 — Filtro eval RCE en Semantic Kernel laboratorio Un laboratorio autocontenido que reproduce CVE-2026-26030 : ejecución remota de código inyectable mediante el filtro de búsqueda en el almacén de vectores en memoria de Microsoft Semantic Kernel Python, Solo laboratorio ético. Aislad...
batch_jailbreak
Repositorio oficial de Safety in Batches? Understanding and Mitigating Safety Failures in Batch Prompting Kihyun Kim, Hee-Seon Kim, Wonjun Lee, Changick Kim Korea Advanced Institute of Science and Technology KAIST Noticias 2026.09 ¡Nuestro artículo ha sido aceptado en AACL-IJCNLP 2026 Main! 🎉...
ContextHound
ContextHound Herramienta de análisis estático que escanea tu base de código en busca de vulnerabilidades de inyección de prompts en LLM y de seguridad multimodal. Se ejecuta sin conexión, no requiere llamadas a API. El ecosistema ContextHound ContextHound está disponible en todo tu flujo de traba...
PT-2026-106004
The convert playwright script prompt in the k6 MCP server accepts a file path as its playwright script argument. Paths given in the documented '@'-prefixed form are restricted to the server's current working directory, but a bare path is resolved by a separate undocumented code path that applies ...
PT-2026-106262
Name of the Vulnerable Software and Affected Versions vLLM versions prior to 0.30.0 Description In vLLM, Harmony tool continuations submitted through the "POST /v1/responses" endpoint rebuild the next-turn engine input without preserving the cache salt value. This causes the continuation prefix t...
RAISED: Self-Distillation for Robustness to Prompt Injection in LLM Agents
Tool-using language-model agents are vulnerable to indirect prompt injection because they must act on untrusted external content. Existing training-time defenses can reduce attack success rates, but often at the cost of general capabilities. We show that training-based defenses induce substantial...
Towards a Unified Misuse Monitoring Benchmark
LLM agents increasingly act in multi-actor environments, exposing them to misuse from multiple sources: decomposition attacks, where a harmful request is split into innocuous sub-requests, and prompt injection attacks, where a compromised tool delivers a malicious instruction. Existing evaluation...
mythic_ornn
Mythic Ornn LLM-driven generator for Mythic Agents, Payload-Type and C2 Profiles. Mythic agents are tedious to write by hand: every one re-implements the same wire protocol, payload-type scaffolding, and C2 plumbing. Mythic Ornn skips the boilerplate by turning the official Mythic developer docs...
CVE-2026-24055-OAuth-Langfuse
CVE-2026-24055 — Instalación de Slack OAuth sin autenticación en Langfuse Proyecto del curso Seguridad Web y de Aplicaciones — Grupo 06, clase NT213.Q21.ANTT 📌 Resumen El proyecto reproduce una vulnerabilidad de seguridad real ya publicada CVE-2026-24055 en la plataforma Langfuse — un endpoint...
CVE-2026-78906-ChatGPT-Prompt-Injection
CVE-2026-78906 — PromptGhost API de ChatGPT — Inyección de Prompts y Exfiltración de Memoria Campo| Valor ---|--- Severidad| 8.7 Alta Vector| Red Versiones Afectadas| API v1.0 – v1.3 Descubierto Por| 𝕍𝕠𝕤𝕤🥷 Descripción Una vulnerabilidad crítica en la API de ChatGPT de OpenAI permite a un atacante...
claude-project-scanner
Escáner de Proyectos Claude Skill de Claude Code de solo lectura que escanea proyectos de terceros en busca de riesgos de seguridad conocidos antes de que los abras. Detecta los patrones de ataque documentados en: CVE-2025-59536 inyección de Lifecycle Hooks de Claude Code CVE-2025-61260 inyección...
rebuff
Rebuff.ai Detector de inyección de prompts con auto-endurecimiento Rebuff está diseñado para proteger aplicaciones de IA contra ataques de inyección de prompts PI mediante una defensa multicapa. Playground • Discord • Características • Instalación • Primeros pasos • Autohospedaje • Contribuciones...
agentic-dm-gateway
Agentic DM Gateway Security control plane for LLM agents over private chat typically Discord DMs. It sits in front of your agent. It decides who may talk, whether the session is unlocked, whether the process is paused, and whether this message is safe enough to forward. Your model and tools stay...
prompt-injection-email-samples
Prompt-Injection Email Samples A small, open test set of emails for checking whether your email security controls and AI mailbox assistants handle indirect prompt injection : instructions hidden in an email that try to hijack an AI assistant when it reads, summarizes or acts on that email. The se...
cve-2026-54316-lab
CVE-2026-54316 — Laboratorio de exfiltración de Claude Code vía WebFetch y HuggingFace Un laboratorio autocontenido y desechable que reproduce GHSA-fg94-h982-f3mm / CVE-2026-54316: Claude Code preaprobó huggingface.co como un nombre de host simple para la herramienta WebFetch, por lo que cualquie...
llm-prompt-injection-resources
llm-prompt-injection-resources Una colección seleccionada de recursos para aprender e investigar sobre ataques de inyección de prompts en LLM, defensas y seguridad. Donar Apoya el mantenimiento de este proyecto con PayPal o escaneando el código QR a continuación...
Blocking at the Boundary: Auditing Long-Horizon Agents against Staged Prompt Injection
Long-horizon agents consume external content, invoke tools, and modify persistent state. Indirect prompt injection can exploit task-specific context, propagate across causally connected stages, and alter a consequential action while the workflow continues; we term this staged prompt injection. We...
Can CaMeLs Talk? Securing Multi-Agent Systems against Indirect Prompt Injection Attacks
Indirect prompt injection attacks - malicious instructions embedded in content processed by large language models - remain a major obstacle to safely deploying tool-using agents. CaMeL Debenedetti et al., 2025 mitigates this threat for an individual agent by separating trusted control flow from...
Hidden Risks of Jev: An Empirical Study of Security, Privacy, and Dual Use
Jev turns natural-language questions into typed answers and probabilities with low latency and cost, enabling applications to route requests and select tools. While this interface allows Jev to integrate naturally into application workflows as a decision layer, the security and privacy implicatio...