2974 matches found
CVE-2026-1731
CVE-2026-1731 BeyondTrust Remote Support Pre-Auth RCE PoC !WARNING This script is intended for educational and research purposes only. Do not use it against systems without explicit permission. Unauthorized access or testing is illegal and unethical. Read the full DISCLAIMER before using this...
n8n-CVE-2025-68613-exploit
n8n - CVE-2025-68613: Improper Control of Dynamically-Managed Code Resources Vulnerability n8n contains a critical Arbitrary Code Execution vulnerability in its workflow expression evaluation system. Under certain conditions, expressions supplied by authenticated users during workflow configurati...
giskard-oss
Evals, Red Teaming and Test Generation for Agentic Systems Modular, Lightweight, Dynamic and Async-first...
bloom
Bloom: Evaluaciones Automatizadas de Comportamiento para LLMs !IMPORTANT Bloom tiene un nuevo hogar. Ahora es desarrollado y mantenido por Meridian Labs y vive en meridianlabs-ai.github.io/petribloom — todas las nuevas funcionalidades y correcciones llegarán allí. Este repositorio está congelado ...
WAFEC
OWASP WAFEC Project Document Repository OWASP WAF Evaluation Criteria Project The Web Application Firewall Evaluation Criteria WAFEC provides interested partied, including users, vendors and 3rd party evaluators with a tool to learn about Web Application Firewalls WAFs and evaluate the suitabilit...
pdfjs_disable_eval
pdfjsdisableeval Module for disabling JavaScript evaluation in PDF.js. This works around the missing fix in Odoo 14.0 for CVE-2024-4367...
CTFTiny
CTFTiny: Lite Benchmarking Offensive Cyber Skills in Large Language Models This is the official repository for CTFTiny from "Towards Effective Offensive Security LLM Agents: Hyperparameter Tuning, LLM as a Judge, and a Lightweight CTF Benchmark" AAAI'26 paper. For CTFJudge, please refer to CTFJud...
MalEval
MalEval Article: Is “Knowing It’s Malicious” Enough? Evaluating LLMs for Fine-Grained Malware Behavior Auditing Article DOI: 10.1145/3832187 MalEval is a framework for evaluating Android malware behavior reports generated by large language models. The code in this repository implements two...
buyer-eval-skill
Buyer Eval — El Rotten Tomatoes del software B2B Un skill gratuito y de código abierto para Claude que evalúa proveedores B2B conversando con sus agentes de IA, contrastando cada afirmación con fuentes independientes y puntuando lo que está verificado frente a lo que es solo marketing optimista...
promptfoo
Promptfoo: evaluaciones de LLM y red teaming promptfoo es una CLI y biblioteca para evaluar y hacer red teaming de aplicaciones LLM. Abandona el enfoque de prueba y error: comienza a lanzar aplicaciones de IA seguras y confiables. Sitio web · Primeros pasos · Red Teaming · Documentación ·...
reverse-captcha-eval
Reverse CAPTCHA: Evaluación de la susceptibilidad de los LLM a la inyección invisible de instrucciones Unicode Un marco de evaluación que prueba si los modelos de lenguaje grandes siguen instrucciones codificadas en Unicode invisible incrustadas en texto que de otro modo parece normal. Mientras q...
santamon
Santamon Lightweight macOS detection sidecar for Santa that evaluates Endpoint Security telemetry locally with CEL rules and forwards only matched detection signals to a backend server. Experimental. Built for home labs and small fleets. Early release – expect bugs and API changes. What It Does...
fuzzer-testing
fuzzer-testing Imágenes Docker estandarizadas para evaluar fuzzers de código abierto, completamente integradas con los corpora para una fácil experimentación. Fuzzers compatibles / incluidos AFL github.com/google/AFL AFLplusplus github.com/AFLplusplus Honggfuzz github.com/google/honggfuzzz QSYM...
EuConform
EuConform 🇪🇺 Open-Source Evidence Toolkit For AI Compliance Open evidence format • Local bias evaluation • Schema validation • CycloneDX interoperability Offline-first • Privacy-preserving • Reusable artifacts • WCAG 2.2 AA accessible EuConform defines an open evidence format for AI compliance and...
ossem-power-up
OSSEM Una herramienta para evaluar la calidad de los datos, construida sobre el increíble proyecto OSSEM. Misión Responder a la pregunta: Quiero empezar a cazar técnicas ATT&CK, ¿qué fuentes de registro y eventos son más adecuados? Crear transparencia sobre las fortalezas y debilidades de tus...
JSONPath Plus < 10.3.0 - Remote Code Execution
Versions of the package jsonpath-plus before 10.3.0 are vulnerable to Remote Code Execution RCE due to improper input sanitization. An attacker can execute aribitrary code on the system by exploiting the unsafe default usage of eval='safe' mode. Note: This is caused by an incomplete fix for...
CVE-2025-54100
CVE-2025-54100 Powershell's curl uses Invoke-WebRequest underneath. root@kitploit: PS C:\Users\melih curl cmdlet Invoke-WebRequest at command pipeline position 1 Supply values for the following parameters: Uri:...
galer
galer root@kitploit: /' '/'' | | /'' ' | || | | | '\ ','\ | /' @dwisiswant0 Una herramienta rápida para obtener URLs de atributos HTML mediante rastreo. Inspirada en el tweet de @omespino, que permite extraer valores src, href, url y action evaluando JavaScript a través del Protocolo Chrome...
lava
LAVA: Large Scale Automated Vulnerability Addition Evaluating and improving bug-finding tools is currently difficult due to a shortage of ground truth corpora i.e., software that has known bugs with triggering inputs. LAVA attempts to solve this problem by automatically injecting bugs into...
clusterfuzz
ClusterFuzz ClusterFuzz is a scalable fuzzing infrastructure that finds security and stability issues in software. Google uses ClusterFuzz to fuzz all Google products and as the fuzzing backend for OSS-Fuzz. ClusterFuzz provides many features which help seamlessly integrate fuzzing into a softwar...