4964 matches found
IPI-exposure-signal
IPI Exposure Signal Este es el repositorio de código de nuestro artículo: Your Agentic LLMs Secretly Encode Latent Signals of Indirect Prompt-Injection Exposure. Versión en arXiv y enlace al artículo: https://arxiv.org/abs/2608.02657 Este repositorio implementa el pipeline de sondeo para señales...
firefox-devtools-mcp
Firefox DevTools MCP Model Context Protocol server for automating Firefox via WebDriver BiDi through Selenium WebDriver. Works with Claude Code, Claude Desktop, Cursor, Cline and other MCP clients. Repository: https://github.com/mozilla/firefox-devtools-mcp Note : This MCP server requires a local...
wardgate
Wardgate - AI Agent Security Gateway Wardgate is a security gateway that sits between AI agents and the outside world -- isolating credentials for API calls, isolating SSH keys for remote command execution, and gating command execution in remote environments conclaves. Give your AI agents access ...
write-ups
Informes RCE en Github Desktop v2.9.4 RCE mediante gh run download GitHub CLI Claude Code: ejecución de código sin sandbox mediante inyección de prompt a través de la confusión del worktree de .git — CVE-2026-55607...
CL4R1T4S
CL4R1T4S AI SYSTEMS TRANSPARENCY AND OBSERVABILITY FOR ALL! Full extracted system prompts, guidelines, and tools from OpenAI, Google, Anthropic, xAI, Perplexity, Cursor, Windsurf, Devin, Manus, Replit, and more – virtually all major AI models + agents! 📌 Why This Exists "In order to trust the...
anamorpher
Anamorpher Anamorpher named after anamorphosis is a tool for crafting and visualizing image scaling attacks against multi-modal AI systems. It provides a frontend interface and Python API for generating images that only reveal multi-modal prompt injections when downscaled. Refer to "Weaponizing...
garak
garak, LLM vulnerability scanner Generative AI Red-teaming & Assessment Kit garak checks if an LLM can be made to fail in a way we don't want. garak probes for hallucination, data leakage, prompt injection, misinformation, toxicity generation, jailbreaks, and many other weaknesses. If you know nm...
beelzebub
Beelzebub Marco de ejecución de engaños Beelzebub es un runtime de engaños de código abierto que despliega servicios señuelo adaptativos e impulsados por LLM en los protocolos SSH, HTTP, TCP, TELNET y MCP. Va más allá de los honeypots pasivos al involucrar activamente a los atacantes en...
clawshield-public
ClawShield Proxy de seguridad para agentes de IA. Se sitúa delante de OpenClaw y escanea cada mensaje en busca de inyección de instrucciones, fugas de PII y secretos, antes de que lleguen al modelo o salgan de la red. Incluye 5 agentes de IA especializados, un panel de control integrado y un moto...
fas-judgement-oss
FAS Judgement Prompt Injection Attack Console Test your AI's defenses before someone else does. Install | Game Mode | Leaderboard | Demo Target | Features | Elite | Contributing Why Judgement? Your AI chatbot, API, or agent is probably vulnerable to prompt injection. Most are. The problem is that...
sigma-ai
AgentShield Sigma Rules What is This Repository? This repository contains detection rules that help identify when an AI agent is being attacked or manipulated. Think of it as a library of "threat signatures" -- each rule describes a pattern that, when matched against an agent's log data, signals...
meta-ai-support-prompt
Meta AI Support Assistant System Prompt Extracted system prompt from Meta's AI Support Assistant on June 1, 2026. Files system-prompt.md — Extracted system prompt ⚠️ Disclaimer & Legal Notice Purpose This repository is published strictly for educational and authorized security research purposes...
pmotadeee
🧠 FLATLINE - Consciousness Injection Protocol WARNING: This is not a game. It's a cognitive interface for reality hacking. 🎮 HOW TO PLAY Basic Gameplay 1. Download consciousness modules CSV files from this repository 2. Inject directly into AI interfaces as file uploads 3. Experience accelerated...
better_opts_attacks
¿Puedo tener su atención? Rompiendo defensas de inyección de prompts basadas en fine-tuning mediante ataques conscientes de la arquitectura Este repositorio contiene el código para ejecutar los ataques ASTRA y ASTRA++ que rompen SecAlign++, SecAlign, StruQ. Este repositorio también contiene algun...
promptmap
O O o.-. ¡Humanos, no os resistáis! |/ ,-'-. / /, / /|.-.| / --O-- .--"""/\ /\ | o.o / De Utku Sen /|\ -'--' / /| | - | / | | | \ /| | | ' \ '/ \ ' | ' \ | ' / | ' \ | / | | | ./| /|| | ./| || , | .// / / |-----| || | | || || root@kitploit: promptmap2 es un escáner automatizado de inyección de...
promptmap
O O o.-. Humans, Do Not Resist! |/ ,-'-. / /, / /|.-.| / --O-- .--""" /\ /\ | o.o / Utku Sen's /|\ -'--' / /| | - | / | | | \ /| | | ' \ '/ \ ' | ' \ | ' / | ' \ | / | | | ./| /||| ./|||,| .// / / |-----| || || || || promptmap2 is a an automated prompt injection scanner for custom LLM...
intentshield
IntentShield No filtres lo que tu IA dice. Filtra lo que está a punto de hacer Verificación de intención previa a la ejecución para agentes de IA. Por Qué Existe Esto Los agentes de IA tienen acceso a herramientas. Pueden ejecutar comandos de shell, escribir archivos, navegar por URLs, enviar...
agent-vault
HTTP credential proxy and vault An open-source credential broker by Infisical that sits between your agents and the APIs they call. Agents should not possess credentials. Agent Vault eliminates credential exfiltration risk with brokered access. New here? Thelaunch blog post has the full story...
proxychains-ng
ProxyChains-NG ver 4.17 README ProxyChains is a UNIX program, that hooks network-related libc functions in DYNAMICALLY LINKED programs via a preloaded DLL dlsym, LDPRELOAD and redirects the connections through SOCKS4a/5 or HTTP proxies. It supports TCP only no UDP/ICMP etc. The way it works is...
dataset
🚀 CySecBench: Generative AI-based CyberSecurity-focused Prompt Dataset for Benchmarking Large Language Models 🛡️ The largest and most comprehensive Generative AI-based CyberSecurity-focused Dataset for Benchmarking Large Language Models 🌟 Overview The CySecBench paper offers: 🎯 A cutting-edge...