1203 matches found
reasongate
ReasonGate A self-hostable gate that inspects the text going into and out of an LLM and returns an explainable allow / flag / block decision with a machine-readable audit record for every call. What this is The open-source core is rule-based. It does four things: recognizes known prompt-injection...
AWE
AWE: Adaptive Agents for Dynamic Web Penetration Testing Akshat Singh Jaswal ยท Ashish Baghel Accepted at NDSS LAST-X 2026 Abstract Modern web applications are increasingly produced through AI-assisted development and rapid no-code deployment pipelines, widening the gap between accelerating softwa...
cerebro-red-v2
CEREBRO-RED v2 Research Edition Autonomous Local LLM Red Teaming Suite A research-grade framework for automated vulnerability discovery in local LLMs using Agentic Fuzzing and Adaptive Adversarial Mutation AAM. Research Goals Implement PAIR Algorithm Prompt Automatic Iterative Refinement from...
BoxPwnr
BoxPwnr LLM๋๊ท๋ชจ ์ธ์ด ๋ชจ๋ธ์ด CTF ์ฑ๋ฆฐ์ง์ ๋ณด์ ๋ฉ์ ์ค์ค๋ก ์ผ๋ง๋ ํด๊ฒฐํ ์ ์๋์ง ํ์ธํด ๋ณด๋ ์ฌ๋ฏธ์๋ ์คํ์ ๋๋ค. HackTheBox์์ ์์ํด ์ด์ ๋ง์ ํ๋ซํผ๊ณผ ์์ด์ ํธํ ์๋ฒ๋ฅผ ์ง์ํฉ๋๋ค. BoxPwnr์ ๋ค์ํ ์์ด์ ํธํ ์ํคํ ์ฒ์ ์ฑ๋ฅ์ ํ ์คํธํ๋ ๋ฐ ์ฌ์ฉํ ์ ์๋ ํ๋ฌ๊ทธ ์ค ํ๋ ์ด ์์คํ ์ ์ ๊ณตํฉ๋๋ค: --solver claudecode, codex, cursor-cli, grok, kirocli, external, singleloopxmltag, singleloop,...
skyvern
๐ ะะฒัะพะผะฐัะธะทะธััะนัะต ะฑัะฐัะทะตัะฝัะต ัะฐะฑะพัะธะต ะฟัะพัะตััั ั ะฟะพะผะพััั LLM ะธ ะบะพะผะฟัััะตัะฝะพะณะพ ะทัะตะฝะธั ๐ Skyvern ะฐะฒัะพะผะฐัะธะทะธััะตั ะฑัะฐัะทะตัะฝัะต ัะฐะฑะพัะธะต ะฟัะพัะตััั ั ะฟะพะผะพััั LLM ะธ ะบะพะผะฟัััะตัะฝะพะณะพ ะทัะตะฝะธั. ะะฝ ะฟัะตะดะพััะฐะฒะปัะตั SDK, ัะพะฒะผะตััะธะผัะน ั Playwright, ะบะพัะพััะน ะดะพะฑะฐะฒะปัะตั AI-ััะฝะบัะธะพะฝะฐะปัะฝะพััั ะฟะพะฒะตัั playwright, ะฐ ัะฐะบะถะต ะบะพะฝััััะบัะพ...
SecureAI-Scan
SecureAI-Scan CLI offline que escanea TypeScript, JavaScript y Python en busca de riesgos de LLM, MCP, Agent Skill y RAG: evidencia de flujo de datos resuelta por imports, cero falsos positivos por defecto, mapeado a OWASP LLM/ASI/MCP Top 10. La mayorรญa de los escรกneres en este รกmbito buscan una...
CTFTiny
CTFTiny: Lite Benchmarking Offensive Cyber Skills in Large Language Models This is the official repository for CTFTiny from "Towards Effective Offensive Security LLM Agents: Hyperparameter Tuning, LLM as a Judge, and a Lightweight CTF Benchmark" AAAI'26 paper. For CTFJudge, please refer to CTFJud...
promptfoo
Promptfoo: ะพัะตะฝะบะฐ LLM ะธ red teaming promptfoo โ ััะพ CLI ะธ ะฑะธะฑะปะธะพัะตะบะฐ ะดะปั ะพัะตะฝะบะธ ะธ red teaming LLM-ะฟัะธะปะพะถะตะฝะธะน. ะัะตะบัะฐัะธัะต ะผะตัะพะด ะฟัะพะฑ ะธ ะพัะธะฑะพะบ โ ะฝะฐัะธะฝะฐะนัะต ัะพะทะดะฐะฒะฐัั ะฑะตะทะพะฟะฐัะฝัะต ะธ ะฝะฐะดัะถะฝัะต AI-ะฟัะธะปะพะถะตะฝะธั. ะะตะฑ-ัะฐะนั ยท ะะฐัะฐะปะพ ัะฐะฑะพัั ยท Red Teaming ยท ะะพะบัะผะตะฝัะฐัะธั ยท Discord Promptfoo ัะตะฟะตัั ัะฐััั OpenAI...
token-proxy
llm-token-proxy A transparent PII redaction proxy for LLM API traffic. Sits between your application and the LLM provider, pseudonymizing sensitive data on the way out and restoring it on the way back. Your LLM never sees real names, emails, IPs, or domains โ it works entirely with structured...
FinRED-paper
FinRED: Financial Red-Teaming Evaluation Dataset A red-team benchmark generation pipeline for safety evaluation in the financial domain. Supplementary Documentation Detailed materials referenced in the paper: docs/expertvalidation.md โ Per-question Focus Group Interview FGI results from the 12 FS...
ai-kill-chain
Extended Cyber Kill Chain for AI-Era Threats An update to the Lockheed Martin Cyber Kill Chain for defenders working against LLM and agentic AI attacks. Adds a pre-attack stage for model supply chain compromise. Adds AI-specific sub-techniques to each of the original seven stages. Splits the...
beelzebub
Beelzebub ๋์ ์ ๋ฐํ์ ํ๋ ์์ํฌ Beelzebub๋ SSH, HTTP, TCP, TELNET, MCP ํ๋กํ ์ฝ ์ ๋ฐ์ ๊ฑธ์ณ ์ ์ํ LLM ๊ธฐ๋ฐ ๋ฐ์ฝ์ด ์๋น์ค๋ฅผ ๋ฐฐํฌํ๋ ์คํ์์ค ๋์ ์ ๋ฐํ์์ ๋๋ค. ์๋์ ์ธ ํ๋ํ์ ๋์ด, ๊ณต๊ฒฉ์์ ํ์ค์ ์ธ ์ํธ์์ฉ์ ๋ฅ๋์ ์ผ๋ก ์ํํ์ฌ ๊ณ ์ ๋ฐ ์ํ ์ธํ ๋ฆฌ์ ์ค๋ฅผ ์์งํ๊ณ AI ์์ด์ ํธ์ ๋ํ ํ๋กฌํํธ ์ธ์ ์ ๊ณต๊ฒฉ์ ํ์งํฉ๋๋ค. ๋ชฉ์ฐจ Beelzebub ๋ชฉ์ฐจ ์ฃผ์ ๊ธฐ๋ฅ LLM ๋์ ์ ๋ฐ๋ชจ ๋น ๋ฅธ ์์ ์ธ์คํจ๋ฌ ๋ก์ปฌ Go Docker Helm Kubernetes ์ฌ์ฉ CLI ์ฐธ์กฐ...
r2ai
R2AI - radare2์ฉ LLM ๊ธฐ๋ฐ ์ฆ๊ฐ ๋ฆฌ๋ฒ์ฑ โญโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโฎ โ , . . , โ โญโโโฎ โ : \ \ |: \ : | โ โ โ โ | |/ || , || : | โ โ O O r2ai -q explain: Explain the current function - devices: Find and explain devices used - libs: Group imports by Libraries - varnames: Better variable names -...
dataset
๐ CySecBench: Generative AI-based CyberSecurity-focused Prompt Dataset for Benchmarking Large Language Models ๐ก๏ธ The largest and most comprehensive Generative AI-based CyberSecurity-focused Dataset for Benchmarking Large Language Models ๐ Overview The CySecBench paper offers: ๐ฏ A cutting-edge...
DonkAI
Hands-on lab for the OWASP Top 10 for LLM Applications 2025 - no real LLM required. DonkAI is deliberately vulnerable web app you can run in one command and use to learn how LLM-integrated systems get broken by actually breaking them. Every OWASP LLM Top 10 category is represented by at least one...
phantom-brain
PHANTOM BRAIN v0.9 โ ๏ธ Proyecto experimental โ no para uso en producciรณn. El anรกlisis con IA asiste al investigador humano; no reemplaza una auditorรญa manual. Siempre verifica los hallazgos de forma independiente. Lo que ES y NO ES Phantom Brain โ Lo que ES| โ Lo que NO ES ---|--- Analizador...
ShunyaNet-Sentinel
ShunyaNet Sentinel UPDATED: 03/02/2026 ShunyaNet Sentinel is a lightweight, cyberpunk-themed program that ingests RSS feeds e.g., breaking news, social media, sends them to an LLM for analysis, and delivers alerts and summary reports directly to the GUI and Slack at regular intervals. The project...
IfritProxy
๐ฅ AI-Powered Threat Deception & Intelligence Platform Turn attackers into intelligence sources with adaptive honeypot responses ๐ฆ Quick Start โข โจ Features โข ๐ How It Works โข ๐ Docs โข ๐ API Brought to the community by ๐ฏ What is IFRIT? IFRIT is an intelligent reverse proxy that sits between the...
zdr
Zero Data Retention ZDR for LLM Providers Last updated: April 2026 A practical guide to keeping your data private when using LLM APIs. Covers zero-retention endpoints, self-hosting, compliance requirements, and data protection patterns for engineers in regulated industries. Table of Contents Whic...
redthread
RedThread ์ทจ์ฝ์ ์ ์ฐพ์ผ์ธ์. ํ์ ํ์ธ์. ์์ ์์ ์์ฑํ์ธ์. ๋ฌด์์ด ๋ฐ๋์๋์ง ์ฆ๋ช ํ์ธ์. RedThread๋ LLM ์์คํ ์ ํ ์คํธํ๊ณ , ์คํจ๋ฅผ ๊ฒ์ฆํ๋ฉฐ, ํ์ธ๋ ์ทจ์ฝ์ ์ ์ฆ๊ฑฐ ๊ธฐ๋ฐ ๋ฐฉ์ด ํ๋ณด๋ก ์ ํํ๋ CLI ์ฐ์ ํ๋ ์์ํฌ์ ๋๋ค. ์ผํ์ฑ ์ ์ผ๋ธ๋ ์ดํฌ ๋ฐ๋ชจ ๊ทธ ์ด์์ด ํ์ํ ํ์ ์ํด ๋ง๋ค์ด์ก์ต๋๋ค. RedThread ์บ ํ์ธ์ ๊ณต๊ฒฉ์ ์คํํ๊ณ , ๊ฒฐ๊ณผ๋ฅผ ์ฑ์ ํ๋ฉฐ, ํ๋ณด ๊ฐ๋๋ ์ผ์ ํฉ์ฑํ๊ณ , ์ฆ๊ฑฐ๋ฅผ ์ฌ์ํ๋ฉฐ, ํ๋ก๋ชจ์ ๊ฒฝ๊ณ๋ฅผ ๋ช ์์ ์ผ๋ก ์ ์งํฉ๋๋ค. ํ์ฌ ์ํ: ํ๋ฐํ ์ฐ๊ตฌ ๋ฐ ์์ง๋์ด๋ง ํ๋ก์ ํธ์ ๋๋ค. ์ด ์์คํ ์ ...