1654 matches found
GHSA-763M-79HH-57F2 vulnerabilities
Vulnerabilities for packages: open-webui...
promptfoo v0.123.0
Promptfoo: LLM evals & red teaming promptfoo is a CLI and library for evaluating and red-teaming LLM-based apps. Stop using trial-and-error... start shipping secure, reliable agents Website · Getting Started · Red Teaming · Documentation · Discord Promptfoo is now part of OpenAI. Promptfoo remain...
PT-2026-89869
Name of the Vulnerable Software and Affected Versions Headroom versions prior to 0.35.0 Description The Headroom WebSocket server fails to validate the Origin header of incoming client WebSocket requests before forwarding them to the upstream server. This allows malicious WebSocket clients, which...
CVE-2026-71416: Improper Authentication
Headroom compresses data before the data reaches a large language model. Prior to version 0.35.0, the Headroom WebSocket server does not validate the Origin header of incoming client WebSocket requests before forwarding the request to the upstream server, allowing malicious WebSocket clients to...
IntentFuzz: A Protocol-Aware Fuzzer for Automated Invariant Violation Detection in Intent-Based Cross-Chain Bridges
Cross-chain bridges move value between blockchains. Intent-based bridges are a variant where a solver fulfills a user's declared outcome and an off-chain settlement layer later reconciles the fill against the deposit. Existing smart-contract fuzzers and static analyzers only flag known-bad code...
Hieronym: Leveraging Hierarchical Multi-Source Information for Function Renaming in Stripped Binary
Function renaming in stripped binaries can substantially assist reverse engineers by improving code readability, yet it is a challenging task. The difficulty stems from the need to accurately capture function semantics from low-level binary code across diverse instruction sets, architectures, and...
What Is the Difference between Me and You? Benchmarking the Quality Gap between Human-Written and AI-Generated Code
AI coding assistants are becoming co-authors of production software, yet their evaluation centers on functional correctness, leaving open whether their code differs from human code in the quality dimensions dominating lifecycle cost. We compare human-written and AI-generated code at scale: 787,56...
BadEngram: Backdoor Attack on Gated Memory Components in LLMs
To expand open-weight models' capacity without proportionally increasing computation, recent language models incorporate gated parametric memories that retrieve learned values and inject them into intermediate representations. Despite these efficiency benefits, such modules create a distinct atta...
LLMVul: A Vulnerability-Labeled Dataset of LLM-Generated C/C++ Functions from Real Production Repositories
Large language models LLMs are increasingly used to generate and assist with software development, yet existing vulnerability datasets largely focus on human-written code or controlled prompting environments. This limits the ability to study security weaknesses in LLM-generated code as it appears...
garak v0.17.0
garak, LLM vulnerability scanner Generative AI Red-teaming & Assessment Kit garak checks if an LLM can be made to fail in a way we don't want. garak probes for hallucination, data leakage, prompt injection, misinformation, toxicity generation, jailbreaks, and many other weaknesses. If you know nm...
praisonaiagents: ast_grep_rewrite rewrites arbitrary files without the @require_approval gate enforced on every sibling mutation tool
Target: PraisonAI MervinPraison/PraisonAIAffected component: praisonaiagents/tools/astgreptool.py — astgreprewriteAffected versions: master at ce97667156a116c50b4a3d1aa21e09f048903fda; reproduced against the current praisonaiagents PyPI release praisonaiagents = 1.6.52. SummaryTools in...
Cognee allows non-superusers to overwrite global LLM configuration
Cognee before 1.2.0 contains an improper access control vulnerability that allows unauthenticated attackers to overwrite the global LLM provider configuration by self-registering an account and calling the settings endpoint, which performs no admin or superuser check. Attackers can redirect all L...
PYSEC-2026-3816 Cognee allows non-superusers to overwrite global LLM configuration
Cognee before 1.2.0 contains an improper access control vulnerability that allows unauthenticated attackers to overwrite the global LLM provider configuration by self-registering an account and calling the settings endpoint, which performs no admin or superuser check. Attackers can redirect all L...
Stateless, but not forgetful: How LLMs can whisper to their future selves
Based on the research paper: “Stateless Yet Not Forgetful: Implicit Memory as a Hidden Channel in LLMs” published at the4th IEEE Conference on Secure and Trustworthy Machine Learning SaTML, 2026...
BlueSTAR: Tiered Agentic Architecture for Autonomous Cyber Defense
Cyber attacks are increasingly automated, narrowing the time available for human analysts to detect, reason about, and respond to intrusions. Large language models LLMs offer a promising foundation for autonomous cyber defense because they can correlate heterogeneous evidence and reason about...
mulot
mulot -4285F4?logo=googlechrome&logoColor=white Agentic AI web pentester that drives a browser. An open-weights LLM GLM-5.2, Gemma or Qwen drives a real headless Chromium through a Burp-style toolkit and works a target the way a human pentester would. No frontier model, no agent running inside a...
Infostealer Logs Expose Replayable AI Tokens That Can Bypass MFA
Cybercriminals are hijacking artificial intelligence AI user accounts via information stealer logs to create "stolen keys" that grant illicit access to tools from model providers like Google, Anthropic, and others. Information stealers like Lumma Stealer or Vidar are equipped to harvest a wide...
CS-Guard: Benchmarking LLM Guardrails for Code Generation Security
Large language models LLMs have been ex- ploited to generate malware, but the effective- ness of guardrails for code generation secu- rity remains unclear. We introduce CS-Guard, the first benchmark to systematically evalu- ate guardrails for code generation security. It covers 1 text-to-code...
PrivAudit: A Dual-Lens Auditing Framework for Website Privacy Practices under the CCPA
Five years after the enforcement of the California Consumer Privacy Act CCPA, understanding how website privacy practices evolve at scale in response to regulation remains a key challenge for both researchers and regulators. Prior work and regulatory efforts have focused on manual and case-specif...
AgentAudit: An Open, Extensible Framework for Full-Lifecycle Trust Evaluation of AI Agents
Existing evaluation frameworks mostly assess only one part of AI agents, such as task completion AgentBench or security robustness AgentDojo, ASB, rather than the complete pipeline of planning, tool selection, tool execution, memory and reasoning. Failures can occur at any stage, yet existing...