850 matches found
pike-agent
Agente Pike pike-agent registra y analiza cómo se comportan los programas en Linux. Rastrea la actividad de un programa, la indexa en una base de datos y te permite chatear con un agente LLM al respecto en una TUI. Ejemplo de indicaciones: Diagnóstico de fallos: This program crashed with a bus...
CTFTiny
CTFTiny: Lite Benchmarking Offensive Cyber Skills in Large Language Models This is the official repository for CTFTiny from "Towards Effective Offensive Security LLM Agents: Hyperparameter Tuning, LLM as a Judge, and a Lightweight CTF Benchmark" AAAI'26 paper. For CTFJudge, please refer to CTFJud...
vuln-scanner
vuln-scanner An automated vulnerability assessment platform that orchestrates 86 open-source security tools , aggregates and deduplicates findings, runs an optional OpenAI-compatible LLM analysis layer for triage, clustering, and remediation, generates proof-of-concept scripts , and produces...
honeyprompt
honeyprompt Presentamos honeyprompt, un framework de decepción basado en LLM creado por/para desarrolladores web. El proyecto personal de @alectrocute. Compatible con todos los principales proveedores de LLM en la nube y locales. SSH, HTTP, TLS, TCP, telnet y más. Se distribuye como un contenedor...
AWE
AWE: Adaptive Agents for Dynamic Web Penetration Testing Akshat Singh Jaswal · Ashish Baghel Accepted at NDSS LAST-X 2026 Abstract Modern web applications are increasingly produced through AI-assisted development and rapid no-code deployment pipelines, widening the gap between accelerating softwa...
DonkAI
Hands-on lab for the OWASP Top 10 for LLM Applications 2025 - no real LLM required. DonkAI is deliberately vulnerable web app you can run in one command and use to learn how LLM-integrated systems get broken by actually breaking them. Every OWASP LLM Top 10 category is represented by at least one...
zdr
Retención Cero de Datos ZDR para Proveedores de LLM Última actualización: abril de 2026 Una guía práctica para mantener tus datos privados al usar APIs de LLM. Cubre endpoints de retención cero, auto-alojamiento, requisitos de cumplimiento y patrones de protección de datos para ingenieros en...
phantom-brain
PHANTOM BRAIN v0.9 Offline AI-powered pentesting analysis tool with real hardware integration Local LLM analysis via Ollama for WiFi, Sub-GHz, NFC/RFID and WPA2 captures — no internet required, no cloud APIs, 100% offline. The main LLM runs on a Windows PC. The Raspberry Pi acts as a lightweight...
token-proxy
llm-token-proxy A transparent PII redaction proxy for LLM API traffic. Sits between your application and the LLM provider, pseudonymizing sensitive data on the way out and restoring it on the way back. Your LLM never sees real names, emails, IPs, or domains — it works entirely with structured...
ClawGuard
ClawGuard 🛡️ 中文版 Our Project:https://github.com/SafeAgent-Beihang/clawguard ClawGuard is a security toolkit designed to mitigate risks associated with autonomous agents, such as OpenClaw and other LLM-driven entities. As agents gain more autonomy to execute code, access APIs, and manage files,...
SecureAI-Scan
SecureAI-Scan Offline CLI that scans TypeScript, JavaScript, and Python for LLM, MCP, Agent Skill, and RAG risks — import-resolved dataflow evidence, zero default false positives, mapped to OWASP LLM/ASI/MCP Top 10...
pmotadeee
🧠 FLATLINE - Consciousness Injection Protocol WARNING: This is not a game. It's a cognitive interface for reality hacking. 🎮 HOW TO PLAY Basic Gameplay 1. Download consciousness modules CSV files from this repository 2. Inject directly into AI interfaces as file uploads 3. Experience accelerated...
Aura-State
Aura-State Un framework en Python para construir flujos de trabajo de LLM como máquinas de estado, con verificación formal integrada. root@kitploit: pip install git+https://github.com/munshi007/Aura-State.git Qué es esto La mayoría de los frameworks de LLM te permiten encadenar llamadas a la API ...
redeval
RedEval - LLM Safety Evaluation Framework A comprehensive framework for evaluating the safety of Large Language Models LLMs through systematic attack and refusal testing. RedEval provides a unified, secure, and extensible platform for assessing LLM robustness against adversarial prompts and harmf...
vader
VADER: A Human-Evaluated Benchmark for Vulnerability Assessment, Detection, Explanation, and Remediation Official GitHub Repo:https://github.com/AfterQuery/vader Hugging Face Dataset:https://huggingface.co/datasets/AfterQuery/vader VADER is a human-evaluated benchmark designed to measure how well...
DVAP
DVAP - Damn Vulnerable AI Platform Train. Break. Defend. AI Systems. An open-source platform for AI security training, red/blue teaming, CTF, benchmarking, and research. Runs 100% locally. No cloud, no paid APIs, no data leaves your machine. Prerequisites Docker 24+ and Docker Compose v2 16 GB RA...
GhidraMCP
ghidraMCP ghidraMCP is an Model Context Protocol server for allowing LLMs to autonomously reverse engineer applications. It exposes numerous tools from core Ghidra functionality to MCP clients. https://github.com/user-attachments/assets/36080514-f227-44bd-af84-78e29ee1d7f9 Features MCP Server +...
hermes-osint
Hermes OSINT v3.0 🏛️🧠 The Agentic OSINT Analyst Conversational AI-driven investigations. Natural language. Expert results. 🤖✨ What's New in v3.0 🎉 Hermes 3.0 represents a complete paradigm shift from pipeline-based tool orchestration to a conversational AI-driven investigation platform. Powered by...
ai-kill-chain
Extended Cyber Kill Chain for AI-Era Threats An update to the Lockheed Martin Cyber Kill Chain for defenders working against LLM and agentic AI attacks. Adds a pre-attack stage for model supply chain compromise. Adds AI-specific sub-techniques to each of the original seven stages. Splits the...
ShunyaNet-Sentinel
ShunyaNet Sentinel ACTUALIZADO: 03/02/2026 ShunyaNet Sentinel es un programa ligero con temática cyberpunk que ingiere fuentes RSS por ejemplo, noticias de última hora, redes sociales, las envía a un LLM para su análisis y entrega alertas e informes resumidos directamente a la interfaz gráfica y ...