7902 matches found
Trustworthy RAG: An Evaluation Agent for Detecting Misinformation and Knowledge Poisoning in Generative AI Systems
Retrieval-Augmented Generation RAG grounds Large Language Model LLM outputs in external knowledge, but RAG systems usually trust whatever they retrieve, creating a Security-Reliability Gap: high semantic relevance does not guarantee factual truth. Adversaries exploit this through knowledge...
Structured but Fragile: On the Limits of LLMs in Cybersecurity Decision-Making
Large language models LLMs are increasingly used in cybersecurity workflows, yet it remains unclear whether they can perform structured security reasoning or merely rely on superficial cues and prior knowledge. We study this question in the context of defence selection over attack graphs derived...
Enhancing User Resilience against AI-Augmented Phishing: A Two-Stage Framework for Detection and Personalized Training
The rapid development of artificial intelligence, including agents and deepfake techniques, has accelerated phishing attacks and lowered the threshold for attackers. Modern phishing attacks now blend multiple tactics, including social engineering, URL spoofing, and AI deepfakes enabling adversari...
Vibe Coding and Web Application Security: A Twin-Prompt Study
Large language models increasingly generate complete web applications from natural-language prompts, raising the question of whether explicitly requesting security best practice improves the result. We study six functionally distinct web applications, each generated in two prompt variants that ar...
Workplace Surveillance and Insider Threat Risk Management: Legal Limits and Privacy Harms
Workplace surveillance is used by organizations to protect corporate assets and monitor employee productivity. This research presents two central arguments on workplace surveillance: although surveillance serves legitimate organizational purposes, over-surveillance can violate legal requirements...
Nimux 1.0.6
Nimux is a native command surface for authorized security assessments. It combines network enumeration, credential validation, Active Directory operations, Kerberos workflows, remote execution, file movement, secrets collection, DCSync, GPO operations, database clients, and SOCKS routing into one...
Security Games on Series-Parallel Attack Graphs with Adaptive Attackers
We study security games on attack graphs, where an adaptive attacker seeks to reach a target by sequentially attempting stochastic controls along the current attack frontier, while a defender allocates limited resources across controls to delay compromise. The attacker may choose among...
TGL-APT: Temporal Graph Learning with Graph Distillation for Efficient APT Investigation
Advanced Persistent Threat APT attacks pose a critical challenge to modern systems, as their stealthy, multi-stage nature renders conventional detection methods ineffective. While provenance graphs provide rich behavioral context for attack investigation, attack-relevant evidence is often sparse...
The Claws in Plain Sight: Unauthorized Context Disclosure through LLM Agent Tool Calls
LLM agents routinely construct tool-call arguments from user profiles, conversation history, retrieved documents, and prior tool results. However, legitimate access to contextual information does not imply authorization to transmit that information for every purpose or destination. We present Cla...
A Case-Control Measurement Study of OSINT Source Effectiveness for Critical Infrastructure Defense
Defenders of critical infrastructure CI subscribe to many public open-source intelligence OSINT feeds without an empirical basis for which feeds actually precede attacks. We provide one. Across 54 confirmed CI cyberattacks from 2010 through 2024 spanning twelve named CI sectors plus a cross-secto...
Faraday 5.23.2
Faraday is a tool that introduces a new concept called IPE, or Integrated Penetration-Test Environment. It is a multiuser penetration test IDE designed for distribution, indexation and analysis of the generated data during the process of a security audit. The main purpose of Faraday is to re-use...
Cudy WR3000 Authentication Bypass / Command Injection
Cudy WR3000 routers suffer from authentication bypass vulnerability due to a hard-coded secret and a code injection vulnerability via unsanitized input...
MaliciousSkillBench: A Comprehensive Benchmark for Malicious Agent Skill Detection
Agent Skills extend LLM agents with reusable instruction packages that may also include scripts, resources, and service configuration. This creates a direct distribution channel for malicious behavior, yet existing malicious-Skill datasets are fragmented across sources, artifact formats, evidence...
Securing Filesystems for Confidential Computing
Confidential computing protects applications inside Trusted Execution Environments TEEs, but it leaves storage vulnerable. Even with disk encryption, a malicious cloud provider can roll back, replay, fork, or tamper with disk state, breaking the integrity and freshness guarantees required by...
A Meta-Study on Replication Papers in Usable Security and Privacy
The field of usable security and privacy research is a young and expanding field, which is still developing standards for its research, e.g. regarding replications. We used a mixed-method approach, in order to get a better understanding of the current state of replications in the field of usable...
Joern 4.0.607
Joern is the bug hunter's workbench. With this tool, you can uncover attack surface, sloppy coding practices, and variants of known vulnerabilities using an interactive code analysis shell. Joern supports C, C++, LLVM bitcode, x86 binaries via Ghidra, JVM bytecode via Soot, and Javascript...
Apple Security Advisory 08-18-2026-1
Apple Security Advisory 08-18-2026-1 - Safari 26.6.1 addresses out of bounds access and use-after-free vulnerabilities...
AiXamine: Unified Black-Box Evaluation of Cross-Dimensional Trade-Offs in LLM Safety, Security, and Privacy
The critical failure modes in deployed large language models LLMs are cross-dimensional: a model can score 99.3 in safety alignment while refusing one in three benign queries, or improve across every capability metric while losing 21 points in privacy. Existing evaluation frameworks that assess...
Structural Inference in Undocumented Mobile Databases: A Reproducible Benchmark for Evaluating Agentic Reasoning in Digital Forensics
Agentic large language models are increasingly used in digital forensic analysis, yet their ability to infer relational structure inside undocumented mobile application databases remains poorly understood. In forensic contexts, structurally incorrect inferences can yield results that appear...
Specter 2.7
Specter turns your Flipper Zero into a pocket counter-surveillance bug-sweep for active 13.56 MHz NFC readers - a hidden card skimmer slipped into a payment terminal, a covert reader behind a door panel, a rogue logger taped under a desk. It passively senses the RF carrier that any powered-on...
MATEE: Efficiently Bridging the Semantic Gap in TrustZone Via Arm Pointer Authentication
Trusted Execution Environments TEEs employ hardware-based isolation mechanisms to safeguard the confidentiality and integrity of sensitive code and data. One such prevalent implementation is Arm TrustZone, which partitions the system into the secure and normal non-secure worlds. However, this...
AEGIS: Preventing Cross-Domain Resource Abuse in MCP
The Model Context Protocol MCP is an open source JSON-RPC protocol that standardizes how large language models LLMs interact with external systems through programmatic functions known as tools. Attackers or malicious agents can exploit certain modalities of these MCP tools to degrade the overall...
Beyond End-To-End Success: Diagnosing Failures in Long-Horizon Security LLM Agents
Long-horizon security LLM agents must carry information and decisions across many dependent interactions, where later actions often depend on services, state, or access discovered much earlier. This makes final task success difficult to interpret: an agent may fail before it ever reaches the poin...
The Software Supply Chain As a Market for Lemons: A Multivocal Review of Trust Signal Collapse
Practitioners evaluating open-source dependencies rely on cheap trust signals, e.g., stars, download counts, and contributor activity, as substitutes for direct code inspection, assuming those signals reflect genuine trustworthiness. Prior work has documented individual signal gaming, but the...
Survival Of~The~Stealthiest: Evolving Low-Entropy Ransomware Via~Genetic Algorithms
Traditional ransomware deployment often relies on massive encryption procedure, triggering immediate detection by modern defense systems. This work introduces a paradigm shift in cryptographic attacks by framing ransomware execution as a Search-Based Software Engineering SBSE optimization problem...
From Noise to Signal: Improving Security Log Anomaly Detection Using LLMs with Endpoint-Specific Logs
Existing approaches to anomalous behaviour log detection, such as Wazuh rely primarily on predefined detection rules, while statistical anomaly detection approaches such as OpenSearch identify deviations from previously observed behavioural patterns. Recent research has investigated LLMs for log...
Chameleon: Robust Defense against Tor Website Fingerprinting Via Many-To-Many Traffic Morphing
Website fingerprinting WF attacks can infer users' browsing activities from encrypted Tor traffic by exploiting side-channel features. Although many WF defenses have been proposed, we find that most existing defenses create learnable web trace mapping features. We further show that robustness...
Zeek 8.0.10
Zeek is a powerful network analysis framework that is much different from the typical IDS you may know. While focusing on network security monitoring, Zeek provides a comprehensive platform for more general network traffic analysis as well. Well grounded in more than 15 years of research, Zeek ha...
Zeek 8.2.2
Zeek is a powerful network analysis framework that is much different from the typical IDS you may know. While focusing on network security monitoring, Zeek provides a comprehensive platform for more general network traffic analysis as well. Well grounded in more than 15 years of research, Zeek ha...
Wapiti Web Application Vulnerability Scanner 3.3.2
Wapiti is a web application vulnerability scanner. It will scan the web pages of a deployed web application and will fuzz the URL parameters and forms to find common web vulnerabilities. This is the source code release...
Malformer: A Multi-Modal Malware Detector Using Transformers
Traditional malware detection systems that rely on a single representation of malware often fail to identify novel threats. These representations of malware binaries, also known as modalities, do not provide the models with sufficient information to discriminate among all samples. Additionally,...
angr 9.3.3
angr is an open-source binary analysis platform for Python. It combines both static and dynamic symbolic "concolic" analysis, providing tools to solve a variety of tasks...
CauSec: Unboxing the Causal Drivers of Static Vulnerability Analysis Performance
Static Application Security Testing SAST tools are widely used in both industry and academia. Such tools often make design choices that sacrifice detection to achieve higher performance, i.e., increased precision, decreased runtime, or increased scalability. These design choices rely on certain...
IriSig-Spoof: A Real-World Benchmark for Time-Robust Satellite RF Fingerprinting and Spoofing Detection
Low Earth orbit LEO satellite Internet is becoming critical communications infrastructure, yet its open wireless links remain vulnerable to satellite impersonation and signal spoofing. Radio frequency fingerprinting RFF offers a potential defense by exploiting transmitter-specific hardware...
CTIFoundry: An Agent-Native Corpus Scaffold for Cyber Threat Intelligence
Cyber threat intelligence CTI is increasingly consumed not by human analysts but by LLM agents that compose multi-step investigations at query time. The harness side of this shift has matured rapidly planning loops, tool protocols, context management, but the corpus side has not: threat reports a...
HARP: Hierarchical Adaptive Ranking with Preference-Adaptive Fusion for Query-Based CVE Prioritization
Vulnerability prioritization is inherently preference dependent, since the same CVE can receive different remediation priority under different operational preference scenarios. Existing scoring systems and ranking methods typically assume a fixed criterion. In practice, organizations already...
Improving LLM-Based SSH Honeypots through Prompting and Fine-Tuning
LLM-based SSH honeypots often use closed cloud LLMs because they give strong shell realism, but cloud models create deployment problems. These include no stable versioning, provider-side changes, attacker-driven cost, and model decommissioning. Local open-weight models avoid these problems, but...
Joern 4.0.606
Joern is the bug hunter's workbench. With this tool, you can uncover attack surface, sloppy coding practices, and variants of known vulnerabilities using an interactive code analysis shell. Joern supports C, C++, LLVM bitcode, x86 binaries via Ghidra, JVM bytecode via Soot, and Javascript...
ldis Lua Disassembler 20260818
This is ldis, a disassembler for lua binary files. It works for binaries compiled by the 5.4 and 5.5 branches of lua, discovering the compiler used to compile into the opcodes, and delivers the assembly opcodes and data in an easy to read form. It is similar to the output produced by luac -l -l,...
SiNMULI: Novel Signed Network Approach for Malicious URL Identification
In today's era of rapid advancements in artificial intelligence, computer security and online safeguarding measures have undergone significant improvements. However, malicious websites continue to facilitate the spread of phishing schemes, fraudulent activities and unsolicited communications...
Autonomous Cyber Defense in Connected Vehicles: A Multi-Agent Approach to V2X Security
A connected vehicle has roughly 100 milliseconds to decide whether an incoming Basic Safety Message is real or fabricated. If a false emergency braking alert reaches the planning pipeline in time, the car brakes - a safety failure triggered by a security failure. Existing intrusion detection...
Agentic AI for Safety-Critical Multi-Drone Systems: Challenges and Opportunities
Multi-drone systems are increasingly positioned for safety-critical missions such as search and rescue SAR and critical infrastructure monitoring. Yet, real-world adoption remains constrained not only by autonomy performance, but by the difficulty of integrating agentic behavior into professional...
SUSE Security Advisory - SUSE-SU-2026:3634-1
SUSE Security Advisory - This update for rsync fixes the following issues...
Packet Sender 8.11.1
Packet Sender is an open source utility to allow sending and receiving TCP, UDP, and SSL encrypted TCP packets as well as HTTP/HTTPS requests and panel generation. The mainline branch officially supports Windows, Mac, and Desktop Linux with Qt. Other places may recompile and redistribute Packet...
Achievement Unlocked: Let'S Get Hacked! an Empirical Study of Cybercrime in the Video Gaming Ecosystem
The ubiquity of the video game industry and its large user base have transformed video games into complex social and economic ecosystems. Unfortunately, this growing popularity also attracts cybercriminals who deliberately exploit game-specific mechanisms to target players. Despite this growing...
Ubuntu Security Notice USN-8640-1
Ubuntu Security Notice 8640-1 - It was discovered that Engrampa incorrectly handled symbolic links when extracting certain archives. An attacker could possibly use this issue to write arbitrary files and execute arbitrary code...
Red Hat Security Advisory 2026-55855-03
Red Hat Security Advisory 2026-55855-03 - An update for libssh is now available for Red Hat Enterprise Linux 10. Issues addressed include buffer overflow, bypass, denial of service, information leakage, and use-after-free vulnerabilities...
SureForms 2.1.2 CSV Injection
SureForms version 2.1.2 and below contains an unauthenticated CSV injection vulnerability. A remote attacker can submit a specially crafted form field containing an encoded spreadsheet formula that is subsequently included in exported CSV data. When an administrator exports the submissions and...
PostgreSQL Remote Code Execution
PostgreSQL versions 14.0 through 14.23, 15.0 through 15.18, 16.0 through 16.14, 17.0 through 17.10, and 18.0 through 18.5 contain a heap buffer overflow in tochar when processing timezone abbreviations. An authenticated remote attacker able to execute SQL statements can exploit the vulnerability ...
Red Hat Security Advisory 2026-55764-03
Red Hat Security Advisory 2026-55764-03 - An update for kernel is now available for Red Hat Enterprise Linux 8. Issues addressed include code execution and out of bounds write vulnerabilities...