7902 matches found
Efficient Hardware Information-Flow Tracking for Pre-Silicon Security Testing
Register-Transfer Level RTL simulation is widely used to test hardware before it is fabricated. To allow testing for security related information flow properties, such as confidentiality and integrity, taint logic can be automatically added to the design to track how information flows through it...
Skynet: Workflow-Level Anomaly Detection for Agentic AI Via Semantic and Structural Modeling
Agentic AI systems execute complex tasks through long-horizon workflows of planning, tool use, and multi-agent coordination. Task failures in these systems often originate from a single step, such as an injected prompt or a flawed plan, and are then amplified through downstream dependencies as th...
Characterizing Contention-Induced Reliability Collapse in KV-Cache Timing Side Channels for Multi-Tenant LLM Serving
Shared key--value KV cache reuse improves large language model LLM serving, but it can also create a timing side channel that reveals whether a prefix is already cached. Previous work shows that such attacks are possible, but their reliability under realistic multi-tenant contention is less...
A Queryable Graph-Based Security Analysis Framework for O-RAN
The Open Radio Access Network O-RAN replaces vendor-locked RANs with a modular and interoperable architecture that fosters competition and accelerates innovation. With this openness comes increased complexity and a larger attack surface, making security a critical concern. Today, assessing O-RAN...
Booz Allen Cyber Weapon Index
The new Booz Allen Cyber Weapon Index CWI confirms that autonomous offensive cyber has arrived. A leading frontier AI model can now independently execute the full cyber kill chain against a real network—crossing a critical threshold from AI-assisted hacking to autonomous cyber operations. This is...
Step-by-Step Malware Development: Evading EDR from Loaders to the Kernel
These are the workshop materials for "Step-by-Step Malware Development: Evading EDR from Loaders to the Kernel", presented at DEF CON 34 and BSidesLV 2026. This workshop explores custom malware development, EDR Architecture and Evasion, C2 customization, and kernel-level techniques using Elastic...
The History Is the Detector: Executing CVE Patch History, End-To-End
Public vulnerability databases collect rich information about known software flaws, including their weakness types, affected components, and related patches. Fixing commits provide the exact code changes that removed these flaws. While these records capture why the original code was unsafe, they...
Governing Bring Your Own AI: A Parameterized Maturity Model
Employees are increasingly using personally owned generative AI tools such as ChatGPT, Gemini, and Claude for their daily work. This practice is known as Bring Your Own AI BYOAI, which is a distinct form of Shadow AI in which employee-authenticated personal accounts are used outside of enterprise...
CONTINUITY: Security-Context Contracts for Composable LLM Agent Controls
LLM agent systems increasingly combine provenance tracking, authorization, policy enforcement, protocol adapters, and execution controls. However, individually correct security mechanisms do not necessarily compose into an end-to-end secure system: security-critical context may be dropped, widene...
Android Debug Bridge Remote Code Execution
CVE-2026-78745 affects HiDPT / Weyon HiDPTAndroid devices using the Hi3751V350 and Hi3751V352EDMO platforms. The vulnerability allows a remote attacker to execute arbitrary code through the Android Debug Bridge adbd daemon...
Understanding the Privacy-Preserving Potential of HTTP/2 against Webpage Fingerprinting
Website fingerprinting WF attacks can infer which webpage a user visits from encrypted HTTPS traffic alone, compromising privacy even without decryption. WF defenses commonly shape traffic through noise, padding, delays, or flow splitting, yet they are most often studied from the perspective of...
Injected and Leaked: Actively Inducing Side-Channel Leakage Using Electromagnetic Injection and Hardware Nonlinearity
Electromagnetic EM side-channel leakage and injection are typically treated as distinct physical phenomena, threatening data confidentiality and integrity respectively. This work investigates how EM injection can be used to amplify side-channel leakage that is otherwise infeasible. We introduce a...
Conformal Prediction for Offensive Security
Despite its introduction more than a quarter century ago, Conformal Prediction CP has seen surprisingly few applications to the cyber security world thus far. In particular, we observe that, while CP has been employed as a defensive measure in many recent works, its use for carrying out attacks...
Faraday 5.24.0
Faraday is a tool that introduces a new concept called IPE, or Integrated Penetration-Test Environment. It is a multiuser penetration test IDE designed for distribution, indexation and analysis of the generated data during the process of a security audit. The main purpose of Faraday is to re-use...
TIER: Threat Implicitness Benchmark for Evaluating LLM Safety Behaviors
Current LLM safety benchmarks largely rely on binary metrics, overlooking how models respond to harmful prompts with varying threat implicitness. We introduce TIER, a Threat Implicitness Benchmark for behavioral safety evaluation of LLMs. TIER covers four risk domains and four threat levels, from...
Zimbra CVE-2026-73570 Incident Response Toolkit
Zimbra CVE-2026-73570 Incident Response Toolkit is a collection of detection and evidence-preservation utilities for administrators investigating suspected exploitation of CVE-2026-73570. The toolkit searches Zimbra and mail logs for exploitation indicators, examines JSP persistence locations,...
The Security Feature Location Problem
Software security must be realized through security features such as authentication and encryption, but which features does a system implement, and where? We present security feature location: the task of relating code locations to security features, enabling developers to understand security...
Federated Attack Campaign Detection Via Contrastive Encoding of Threat Indicators in Gradient Updates
Detecting orchestrated cyberattack campaigns that span multiple organizations traditionally requires sharing sensitive telemetry and threat intelligence across institutional boundaries and country borders, a barrier that Federated Learning removes by training shared threat detectors directly on...
Cost-Aware Hierarchical Multi-Agent Ransomware Detection and Family Attribution
Ransomware detection and family attribution require analysis of different modalities because it can use packing, obfuscation, process manipulation and runtime evasion techniques. However, conventional multimodal usually uses all available modalities for every sample resulting in unnecessary...
Joern 4.0.618
Joern is the bug hunter's workbench. With this tool, you can uncover attack surface, sloppy coding practices, and variants of known vulnerabilities using an interactive code analysis shell. Joern supports C, C++, LLVM bitcode, x86 binaries via Ghidra, JVM bytecode via Soot, and Javascript...
Candidate Comparability Before Promotion: Conditional Validation in Adaptive Network Intrusion Detection
Adaptive network intrusion detection systems retrain classifiers after drift alarms, but an alarm detects change; it does not establish that a challenger should replace the deployed incumbent. Promotion is security-relevant because it changes the model responsible for subsequent attack detection,...
Repeat-After-Me: Black-Box Adaptive Visual Prompt Injection
Prompt injection is widely recognized as a major security threat to AI agents that interact with untrusted external data, such as websites, documents, and emails. Prior work has shown that, in the text domain, black-box prompt injection can achieve near-perfect attack success rates ASRs. In the...
Data-Optimized Contingency Screening: A Machine Learning Approach to Power System Security
Ensuring the security of the power system is essential for stability and reliability, especially in the event of disruption. Effective classification of contingency in power systems enables proactive decision-making and mitigates large-scale breakdowns and failures. This study explores the use of...
AlcaTRAz - Anchored Tree-Rule Defense against Jailbreaks
Large language models LLMs are vulnerable to jailbreak attacks that bypass safety alignment through carefully crafted prompts. Many existing defenses require access to model weights or internals, making them difficult to apply to black-box deployments. We propose AlcaTRAz Anchored Tree-Rule defen...
Nebulon Enterprise Simulated Threats for Phishing Research (NEST-Phish): A Synthetic Enterprise Phishing Email Dataset for Behavioral and Machine-Learning Research
Phishing remains one of the most persistent cyber threats, yet publicly shareable datasets for studying phishing in realistic enterprise email settings remain limited. To address this gap, we introduce a synthetic enterprise phishing email dataset built around a fictitious organization, Nebulon...
AI-Assisted Design of a Post-Quantum Cryptographic Accelerator: A Deployed-Silicon Case Study
Post-quantum migration is mandated on published timelines, and silicon that ships with a defect cannot be patched remotely. The standard acceptance gate cannot detect an entire class of ML-DSA defects. Signing resamples until a candidate meets its norm bounds, so the executed path varies with the...
Shifting from Injection to Interaction: Rethinking Web Security in the Age of LLMs and Beyond
Large language models LLMs are becoming integral to web applications and browser agents, transforming online interactions while introducing new attack vectors and reshaping longstanding web vulnerabilities. Classical threats such as cross-site scripting XSS can be amplified through LLM-mediated...
A Blind Trust, the Bloody Thrust: When Attacker-Controlled Hook Updates Steer AI Agent Harnesses Towards Malicious Behaviors
Modern AI agent harnesses expose lifecycle hooks that bind shell commands to runtime events such as session start, tool calls, and file edits. These commands run with host privileges yet ship as lifecycle-hook configuration and may fire at times the LLM never observes. We identify the...
CISA: Communicating under Pressure: Best Practices for Service Providers
Developed by CISA, the Federal Bureau of Investigation, and international partners, this guidance describes how organizations can plan and execute clear, timely, accurate, and audience-appropriate communications during IT and operational technology OT outages. Whether caused by cyber threat actor...
angr 9.3.4
angr is an open-source binary analysis platform for Python. It combines both static and dynamic symbolic "concolic" analysis, providing tools to solve a variety of tasks...
Mind the Gap: Robustness Risks in PII Detection Systems
Personally Identifiable Information PII detection is a foundational component of data protection infrastructure where missed entities constitute direct privacy and security risks. Although modern PII systems report strong performance on standard benchmarks, we show that these evaluations mask...
Security and Privacy in the Musical Metaverse: Threat Analysis and Design Implications
The Musical Metaverse MM introduces immersive, real-time environments for collaborative musical interaction, characterized by ultra-low-latency constraints, continuous multimodal data streams, and heterogeneous devices. These properties create a distinctive security and privacy landscape that...
SENTINEL-RL: Offloading Topological Reasoning from LLM Agents in the Security Operations Center
Large language model LLM agents are increasingly proposed as autonomous SOC analysts, but two limitations make them unreliable at enterprise scale: a finite context window cannot hold a multi-thousand-host authentication graph, and free-form generation offers no guarantee that a recommended...
Beyond the Trust Boundary: A Critical Reassessment of the FIDO2 Threat Model
FIDO2/WebAuthn has been widely deployed as a phishing-resistant authentication scheme. Because FIDO2 relies on public-key cryptography and hardware-backed authenticators, its security is often assumed to be guaranteed by design, provided that the cryptographic implementation is correct. In this...
Practice Makes (Im)Perfect: A Look Back at Benchmarking Practices for Microarchitectural Side-Channel Attacks
Microarchitectural side-channel research has grown at an exceptional pace in recent years, increasing the need for rigorous and meaningful benchmarking. Early attack papers typically relied on indirect proxies, such as covert-channel bandwidth or key-recovery on naive AES and RSA implementations,...
Privacy, Robustness, and Fairness Trade-Offs in Federated Intrusion Detection: Geometric Indistinguishability at the Aggregation Interface
Federated learning enables privacy-conscious collaboration for network intrusion detection without centralizing sensitive traffic data, yet its deployment in operational environments must simultaneously satisfy three competing requirements: formal differential privacy guaranties, tolerance to...
HP Easy Start for macOS 2.16.x Privilege Escalation
HP Easy Start for macOS versions prior to 2.16.7.260722 suffer from multiple broken privilege boundary issues...
An Empirical Analysis of CodeQL False Positives and Query Refinements for Java Vulnerabilities
Static application security testing SAST tools help developers find vulnerabilities before deployment, but false positives create substantial triage effort. We study whether CodeQL false positives in Java security analysis form recurring, explainable patterns that can be reduced by refining the...
Engineered Persuasion: Evaluating Personalized Pretexts in LLM-Generated Spear Phishing
Large language models can insert workplace details into phishing pretexts at low cost, but those details may either support or undermine a message's credibility. We recruited 180 U.S. working adults to evaluate simulated, AI-generated phishing emails in a disclosed survey. The emails used four...
Microsoft UFO 3.0.7 Mobile MCP Missing Authentication Checker
Microsoft UFO versions through 3.0.7 contain a missing authentication vulnerability in the Mobile MCP HTTP services. The affected data collection and mobile action servers can expose ADB-backed Android information and device interaction functionality to unauthenticated network clients when the...
The Native-Signature Boundary in Post-Quantum Distributed Authorization
Post-quantum signature migration poses a distinct systems problem when authorization is distributed among multiple parties. In native threshold signing, the signature algorithm may determine key generation, share state, preprocessing, interaction, combination, refresh, and recovery. Architectures...
CAPTCHAs in the Agentic Era: Solvers That Learn from Every Encounter
Vision-language models VLMs can solve visual CAPTCHAs without task-specific training, but the agents built on them approach every challenge from scratch. For such an agent, the hundredth instance of a familiar puzzle costs as much time and compute as the first. Specialized detectors invert the...
SpiderSapien: Client-Centric Web Crawler and Security Scanner
Black-box web application crawling and scanning play an important role for security testing of web applications. Yet state-of-the-art scanners fall short of addressing key characteristics of a modern web application: its extreme dynamism and interactivity on the client side. This paper identifies...
A Bayesian Correlated Equilibrium for Early Insider-Threat Detection
We model insider threat detection as a dynamic Bayesian game in which a platform coordinates a committee of strategic certifiers to sustain equilibrium among honest users and detect malicious deviations before exfiltration. Certifiers and users operate under a Bayesian Temporal Correlated...
Boundary-Mutation Testing for Pattern-Based Secret Detection: A Rule-Level Method and Cross-Scanner Evaluation
Pattern-based secret scanners are commonly validated with example-based fixtures that fix one variable: the text surrounding a credential. We introduce boundary-mutation testing to vary that context, generating credentials from each rule's own regular expression, embedding them in realistic sourc...
The Implications of Linguistic Illegibility for LLM Security
LLMs are trained to generate natural language. However, various strands of evidence indicate that an LLM's externalized linguistic outputs and mechanistically-extracted linguistic features can be an unreliable lens for understanding internal model computation. We introduce the term "linguistic...
Specter 2.8 NFC Reader / Skimmer Bug-Sweep
Specter turns your Flipper Zero into a pocket counter-surveillance bug-sweep for active 13.56 MHz NFC readers - a hidden card skimmer slipped into a payment terminal, a covert reader behind a door panel, a rogue logger taped under a desk. It passively senses the RF carrier that any powered-on...
American Fuzzy Lop plus plus 5.03c
Google's American Fuzzy Lop is a brute-force fuzzer coupled with an exceedingly simple but rock-solid instrumentation-guided genetic algorithm. afl++ is a superior fork to Google's afl. It has more speed, more and better mutations, more and better instrumentation, custom module support, etc...
Video-Based Palm-Vein Authentication under Challenging Conditions
Palm-vein biometrics are increasingly used for secure, contactless authentication. Yet real-world deployment exposes them to surface noise sweat, dirt, illumination and motion variation, and temperature-driven changes in vascular visibility, which remain underexplored for lack of data captured...
Evaluating ML-Based Intrusion Detection Systems: The Illusion of Model Efficacy
Intrusion Detection has been revolutionized due to the integration of Machine Learning. Improved detection rates, reduced false alarms, and optimized algorithms contribute to the perception of improved systems with optimal accuracy and near-perfect performance, the illusion of model efficacy...