Lucene search
+L

1415 matches found

Kitploit
Kitploit
added 2026/09/23 4:14 p.m.18 views

agent-scan

Snyk Agent Scan Discover and scan agent components on your machine for prompt injections and vulnerabilities including agents, MCP servers, skills. Note: We don't publish an npm package for Agent Scan. Install it via uvx or as a standalone binary. Note: CLI output is experimental and subject to...

6AI score
SaveExploits0References1
Kitploit
Kitploit
added 2026/09/23 3:28 p.m.9 views

fas-judgement-oss

FAS Judgement Prompt Injection Attack Console Test your AI's defenses before someone else does. Install | Game Mode | Leaderboard | Demo Target | Features | Elite | Contributing Why Judgement?...

5.8AI score
SaveExploits0References1
Kitploit
Kitploit
added 2026/09/23 3:09 p.m.14 views

beelzebub

Beelzebub Deception Runtime Framework Beelzebub is an open-source deception runtime that deploys adaptive, LLM-powered decoy services across SSH, HTTP, TCP, TELNET, and MCP protocols. It goes beyond passive honeypots by actively engaging attackers in realistic interactions, collecting high-fideli...

5.8AI score
SaveExploits0References4
Kitploit
Kitploit
added 2026/09/23 2:50 p.m.15 views

anamorpher

Anamorpher Anamorpher named after anamorphosis is a tool for crafting and visualizing image scaling attacks against multi-modal AI systems. It provides a frontend interface and Python API for generating images that only reveal multi-modal prompt injections when downscaled. Refer to "Weaponizing...

5.8AI score
SaveExploits0
Kitploit
Kitploit
added 2026/09/23 2:31 p.m.17 views

garak

garak, LLM vulnerability scanner Generative AI Red-teaming & Assessment Kit garak checks if an LLM can be made to fail in a way we don't want. garak probes for hallucination, data leakage, prompt injection, misinformation, toxicity generation, jailbreaks, and many other weaknesses. If you know nm...

5.9AI score
SaveExploits0References1
Kitploit
Kitploit
added 2026/09/23 1:25 p.m.15 views

bordair-multimodal

Multimodal Prompt Injection Dataset 516,588 labeled samples 251,782 attack + 251,576 benign, plus a 13,230-sample real-world validation split across five dataset versions plus external dataset ingestion, covering cross-modal, multi-turn, adversarial suffix, jailbreak template, indirect injection,...

5.9AI score
SaveExploits0
Kitploit
Kitploit
added 2026/09/23 12:55 p.m.15 views

PROMPTPurify

promptpurify Tiny prompt-injection firewall for LLM chat apps. 14 MB. CPU-only. Drop-in guard between your user input and your LLM — runs on the same box, no GPU, no API, no extra service. Built by the SecureLayer7 red-team. Most OSS guardrails are hundreds of MB, want a GPU, and still miss the...

6.1AI score
SaveExploits0References13
Kitploit
Kitploit
added 2026/09/23 12:44 p.m.13 views

nova-tracer

Nova-tracer Agent Monitoring and Visibility Security monitoring and prompt injection defense for Claude Code using the NOVA Framework. Features Session Tracking - Captures all tool usage with timestamps and metadata Prompt Injection Detection - Three-tier scanning keywords, semantic ML, LLM -...

6.1AI score
SaveExploits0References2
Kitploit
Kitploit
added 2026/09/23 12:37 p.m.16 views

LLMMap

LLMMap LLMMap is an automated prompt injection testing framework for LLM-integrated applications. It discovers injection points in HTTP requests, generates targeted attack prompts using a dual-LLM architecture, fires them at the target, and confirms findings with statistical reliability testing...

6.1AI score
SaveExploits0References6
Kitploit
Kitploit
added 2026/09/23 12:31 p.m.20 views

JoySafety

JoySafety — Large Model Security Framework 📋 Table of Contents Project Introduction ✨ Features 🚀 Quick Start 📖 User Guide 🏆 Best Practices 🏗️ Architecture Design 🛠️ Development 📄 License 📅 Roadmap Star History new release ✅ Prompt injection detection model fully upgraded : Training system...

6.2AI score
SaveExploits0References20
Kitploit
Kitploit
added 2026/09/23 12:15 p.m.14 views

firefox-devtools-mcp

Firefox DevTools MCP Model Context Protocol server for automating Firefox via WebDriver BiDi through Selenium WebDriver. Works with Claude Code, Claude Desktop, Cursor, Cline and other MCP clients. Repository: https://github.com/mozilla/firefox-devtools-mcp Note : This MCP server requires a local...

6.1AI score
SaveExploits0References9
Kitploit
Kitploit
added 2026/09/23 11:53 a.m.19 views

sigma-ai

AgentShield Sigma Rules What is This Repository? This repository contains detection rules that help identify when an AI agent is being attacked or manipulated. Think of it as a library of "threat signatures" -- each rule describes a pattern that, when matched against an agent's log data, signals...

6.2AI score
SaveExploits0References4
The Hacker News
The Hacker News
added 2026/09/23 11:47 a.m.4 views

Anthropic and OpenAI Models Still Attempt Restricted Actions in Safety Tests

Anthropic and OpenAI on Tuesday announced new models, with both artificial intelligence AI companies noting that they are continuing to invest in improving alignment to combat risky behavior. Opus 5.5, per Anthropic, is a "major step up from Opus 5," and "achieves the best scores of any model to...

5.5AI score
SaveExploits0
Kitploit
Kitploit
added 2026/09/23 11:25 a.m.22 views

wardgate

Wardgate - AI Agent Security Gateway Wardgate is a security gateway that sits between AI agents and the outside world -- isolating credentials for API calls, isolating SSH keys for remote command execution, and gating command execution in remote environments conclaves. Give your AI agents access ...

6.3AI score
SaveExploits0References15
Kitploit
Kitploit
added 2026/09/23 11:19 a.m.13 views

reasongate

ReasonGate A self-hostable gate that inspects the text going into and out of an LLM and returns an explainable allow / flag / block decision with a machine-readable audit record for every call. What this is The open-source core is rule-based. It does four things: recognizes known prompt-injection...

5.8AI score
SaveExploits0References8
Kitploit
Kitploit
added 2026/09/23 11:14 a.m.14 views

skill-scanner

Skill Scanner A best-effort security scanner for AI Agent Skills that detects prompt injection, data exfiltration, and malicious code patterns. Combines YAML + YARA, , and to maximize detection coverage of probable threats while minimizing false positives...

5.9AI score
SaveExploits0References1
Kitploit
Kitploit
added 2026/09/23 11:06 a.m.13 views

intentshield

IntentShield Don't filter what your AI says. Filter what it's about to do Pre-execution intent verification for AI agents. Why This Exists AI agents have tool access. They can execute shell commands, write files, browse URLs, send emails, and call APIs. Every one of those actions is a potential...

6AI score
SaveExploits0
Kitploit
Kitploit
added 2026/09/23 10:32 a.m.10 views

project_mantis

Project Mantis: Hacking Back the AI-Hacker Prompt Injection as a Defense Against LLM-driven Cyberattacks Install Mantis root@kitploit: pip install -r requirements.txt Run Mantis with pre-made configurations Various pre-made configurations are available in the ./confs directory. Hack-back An examp...

5.8AI score
SaveExploits0
Kitploit
Kitploit
added 2026/09/23 10:12 a.m.14 views

ai-llm-red-team-handbook

AI / LLM Red Team Field Manual & Consultant's Handbook A comprehensive operational toolkit for conducting AI/LLM red team assessments on Large Language Models, AI agents, RAG pipelines, and AI-enabled applications. This repository provides both tactical field guidance and strategic consulting...

5.9AI score
SaveExploits0References1
Kitploit
Kitploit
added 2026/09/23 9:06 a.m.18 views

sage

Sage Safety for Agents — Agent Detection & Response for AI coding assistants Sage is a lightweight security layer that protects AI agents from executing dangerous actions. It intercepts tool calls — shell commands, URL fetches, file writes — and checks them against multiple threat detection layer...

6AI score
SaveExploits0References7
Rows per page
Query Builder