18 matches found
SecureAI-Scan
SecureAI-Scan CLI offline que escanea TypeScript, JavaScript y Python en busca de riesgos de LLM, MCP, Agent Skill y RAG: evidencia de flujo de datos resuelta por imports, cero falsos positivos por defecto, mapeado a OWASP LLM/ASI/MCP Top 10. La mayoría de los escáneres en este ámbito buscan una...
open-source-llm-scanners
open-source-llm-scanners GitHub上按星标排序的开源LLM安全扫描器列表。不提供深入分析。 最低需要10个星标作为门槛 😁 相关:open-source-web-scanners LLM扫描器 能够发现LLM应用中一系列漏洞的工具。 项目主页| 最近提交| 贡献者| 星标 ---|---|---|--- Promptfoo| | | garak| | | Giskard| | | Purple Llama| | | PyRIT| | | Agentic Security| |...
DonkAI
Laboratorio práctico para el OWASP Top 10 para Aplicaciones LLM 2025 - no se requiere ningún LLM real. DonkAI es una aplicación web deliberadamente vulnerable que puedes ejecutar con un solo comando y usar para aprender cómo se rompen los sistemas integrados con LLM, rompiéndolos de verdad. Cada...
reasongate
ReasonGate A self-hostable gate that inspects the text going into and out of an LLM and returns an explainable allow / flag / block decision with a machine-readable audit record for every call. What this is The open-source core is rule-based. It does four things: recognizes known prompt-injection...
aco-prompt-shield
aco-prompt-shield 🛡️ Detén los ataques de inyección de prompt antes de que lleguen a tu LLM — sin costes de API, funciona completamente en local, se integra en 2 minutos. La inyección de prompt es el riesgo de seguridad 1 para aplicaciones LLM. aco-prompt-shield detecta patrones de jailbreak...
OpenHunterAI
OpenHunterAI Tu equipo rojo de IA local. Razonamiento estilo atacante para la seguridad de aplicaciones web, API y LLM. Inicio rápido · Habilidad de agente · Modelo de evaluación · Arquitectura · Documentación · Historial de estrellas OpenHunterAI reúne el alcance, la actividad de escaneo, los...
LLM-MCP-Security-Field-Guide
🛡️ Guía de Seguridad de IA — Seguridad de LLM y MCP La referencia de seguridad más completa, actualizada y orientada a profesionales para aplicaciones LLM y despliegues del Protocolo de Contexto de Modelo MCP. Cubre CVEs reales, patrones de ataque en vivo, marcos OWASP, herramientas de red team y...
cover
Cover Keep private data, internal infrastructure and secrets out of cloud coding agents without breaking your workflow. Install · Quick start · Policies · Monitoring · Pi / OMP · Security Cover is a bidirectional privacy proxy for AI coding agents. It replaces matched sensitive values locally wit...
llm-prompt-injection-resources
llm-prompt-injection-resources Una colección seleccionada de recursos para aprender e investigar sobre ataques de inyección de prompts en LLM, defensas y seguridad. Donar Apoya el mantenimiento de este proyecto con PayPal o escaneando el código QR a continuación...
bordair-detector
Bordair Detector Un detector de dos etapas para intentos de inyección de prompts y jailbreak en entradas de LLM. Una puerta de regex resuelve la gran mayoría del tráfico sin tocar el modelo; cualquier cosa ambigua pasa a un clasificador DeBERTa-v3 cuantizado que se ejecuta en ONNX. El mismo motor...
llm-security-101
Seguridad de LLM 101 Adentrándonos en el ámbito de la seguridad de LLM: una exploración de herramientas ofensivas y defensivas, revelando sus capacidades actuales. A medida que adoptamos los Modelos de Lenguaje de Gran Tamaño LLM en diversas aplicaciones y funcionalidades, es crucial comprender l...
Risk-Adjusted Harm Scoring for Automated Red Teaming for LLMs in Financial Services
The rapid adoption of large language models LLMs in financial services introduces new operational, regulatory, and security risks. Yet most red-teaming benchmarks remain domain-agnostic and fail to capture failure modes specific to regulated BFSI settings, where harmful behavior can be elicited...
CacheTrap: Injecting Trojans in LLMs without Leaving Any Traces in Inputs or Weights
Adversarial weight perturbation has emerged as a concerning threat to LLMs that either use training privileges or system-level access to inject adversarial corruption in model weights. With the emergence of innovative defensive solutions that place system- and algorithm-level checks and correctio...
CAVGAN: Unifying Jailbreak and Defense of LLMs Via Generative Adversarial Attacks on Their Internal Representations
Security alignment enables the Large Language Model LLM to gain the protection against malicious queries, but various jailbreak attack methods reveal the vulnerability of this security mechanism. Previous studies have isolated LLM jailbreak attacks and defenses. We analyze the security protection...
JavelinGuard: Low-Cost Transformer Architectures for LLM Security
We present JavelinGuard, a suite of low-cost, high-performance model architectures designed for detecting malicious intent in Large Language Model LLM interactions, optimized specifically for production deployment. Recent advances in transformer architectures, including compact BERTDevlin et al...
CVE-2025-23254
NVIDIA TensorRT-LLM for any platform contains a vulnerability in python executor where an attacker may cause a data validation issue by local access to the TRTLLM server. A successful exploit of this vulnerability may lead to code execution, information disclosure and data tampering...
Qualys TotalAI: The Journey from LLM Scanner to Comprehensive AI Security Solution
Embarking on the AI/ML Journey The launch of Qualys TotalAI marks a significant milestone in our journey with AI/ML. It all began in March 2024 when we ventured into the rapidly evolving AI/ML landscape and the emerging LLM ecosystem. Recognizing the potential of these technologies to revolutioni...