7902 matches found
LAAF: A Layered Accountability Architecture Framework for LLM Applications
Large Language Models LLMs operate in hospitals, courtrooms, banks, and public service desks, where fluent, confident outputs are treated as authoritative even when ungrounded or incorrect. When such an output contributes to harm, who is answerable, and through what mechanisms can responsibility ...
EBPF-Based Cybersecurity Mechanisms: A Systematic Literature Review
Extended Berkeley Packet Filter eBPF has emerged as a kernel-level framework enabling dynamic security enforcement in modern operating systems. While eBPF's cybersecurity potential has attracted significant attention, existing work remains fragmented across domains, evaluation methodologies, and...
The Guard That Cried Wolf: How Scary Words Make Agent Guardrails Refuse Legitimate Actions
Agent guardrails are checks that approve or refuse each action before an LLM executes it. Sometimes they refuse requests that are genuinely safe. This over-safety blocks deployment when a guardrail refuses an authorized task. Evaluating over-safety is hard: at the boundary an authorized action...
NethServer WebTop 1.5.6 Cross Site Scripting
NethServer WebTop module versions 1.5.6 and below suffer from two persistent cross site scripting vulnerabilities...
Circuit Discovery Helps Detect LLM Jailbreaking: A Mechanistic Interpretability Study
Despite extensive safety alignment, large language models LLMs remain vulnerable to jailbreak attacks that bypass safeguards to elicit harmful content. While prior work attributes this vulnerability to safety training limitations, the internal mechanisms by which LLMs process adversarial prompts...
Are We Shooting Flies with Cannons? Trade-Off Analysis for AI-Based 5G Intrusion Detection
The increasing adoption of Artificial Intelligence AI in network intrusion detection raises the question of whether complex and computationally expensive models are justified for this task. In this work, we investigate the trade-off between detection performance and computational cost for intrusi...
From Verdict to Diagnosis: Attributable Security Review of Pull Requests
Automated code reviewers are increasingly used as gates on pull requests PRs, yet evaluations measure whether they block a malicious change. A block may be triggered by an unrelated issue rather than the vulnerability that makes the PR unsafe; fixing the reported issue can leave the target defect...
MMJailBench: A Factorized Benchmark for Disentangling Multimodal Jailbreak Vulnerabilities
Multimodal Large Language Models MLLMs are increasingly deployed in real-world applications, yet how different factors shape their jailbreak vulnerabilities remains poorly understood. Existing benchmarks often couple harmful intent, prompt framing, visual semantics, and instruction carrier within...
CodeQL 2.26.4
Discover vulnerabilities across a codebase with CodeQL, an industry-leading semantic code analysis engine. CodeQL lets you query code as though it were data. Write a query to find all variants of a vulnerability, eradicating it forever. Then share your query to help others do the same...
SkillShield: Prompt-Space Security Skills for LLM Coding Agents
A coding agent edits files and executes shell commands with its developer's privileges, allowing malicious requests to translate directly into harmful actions or functional malware. Existing defenses have complementary limitations: weight-level alignment is unavailable to API-only deployers,...
Unsaid, Unsafe? Implicit Security Obligations in LLM-Based RTL Code Generation
Large Language Models LLMs generate register-transfer-level RTL code with rapidly improving functional correctness. Security of LLM-generated code, however, has been studied mainly for software, where flaws can still be patched after deployment. Insecure RTL offers no such remedy once taped out...
An Analysis of the Impact of Psychological Factors and Techniques across Different Types of Social Engineering
Phishing is a well-known social engineering SE type used to trick individuals into revealing personal information or performing desired actions, like downloading and installing malware. Other SE types, like vishing and smishing, have emerged and are increasingly being used. As SE continues to...
A Hybrid Security Framework for Mini-Programs: Visual UI Compliance and Network Risk Assessment
With the continuous development of the WeChat ecosystem, WeChat Mini Programs, due to their advantages of not requiring installation, using little memory, and being ready to use instantly, have seen a surge in user numbers and have now become an indispensable service carrier in mobile internet...
A Self-Evolving Multi-Agent Framework Defense against LLM Jailbreak Attacks
Large language models LLMs remain vulnerable to jailbreak attacks that exploit techniques such as role-playing, obfuscation, code transformation, and multi-step indirection to elicit harmful outputs. As jailbreak strategies keep emerging, defenses have proliferated in an ongoing cat-and-mouse gam...
CISA: Vulnerability Review for Fiscal Years 2024 and 2025
Most compromises do not rely on advanced techniques or cutting-edge tools. Cyber threat actors scan the internet looking for exposed, well-known software vulnerabilities to exploit. Basic security failures enable most compromises and organizations can reduce their risk by addressing these...
CISA: A Tale of Two SOCs - Insights from Two Red Team Assessments
The Cybersecurity and Infrastructure Security Agency CISA conducted simultaneous red team assessments at two organizations and observed different defensive outcomes. In both environments, the red team achieved full domain compromise and accessed sensitive business systems SBSs and cloud resources...
SILK: Closing the Time-Of-Check-To-Time-Of-Use Gap in RoT-Protected AI Systems
Root-of-trust RoT authentication verifies a DNN model at load time, but weights may subsequently traverse DRAM, DMA, interconnect, and prefetch paths before reaching the compute engine. Post-verification tampering along this path can therefore alter the weights actually consumed while leaving the...
How Do LLM Agents Actually Get the Flag? Trace-Level Provenance for Agentic Offensive Security Evaluation
Capture-the-Flag CTF benchmarks are widely used to assess the offensive security capabilities of autonomous language-model agents. Evaluations rely on shallow binary judgments or aggregate scores, overlooking the agent's trajectory to the flag. Consequently actual exploitation is conflated with...
Benchmarking Confidential Computing Performance on NVIDIA Blackwell GPUs
This paper measures the performance impact of running large language model inference and training inside a Trusted Execution Environment TEE on NVIDIA B200 GPUs, using Intel Trust Domain Extensions TDX confidential VMs together with NVIDIA Confidential Computing CC on Blackwell GPUs. The...
Adversarial Training of Linear Models under Stealthy Attacks
Predictive models are widely used in many fields, but are vulnerable to false data injection attacks. To address this, detection schemes and adversarial training have been proposed, but such approaches lack guarantees against stealthy attacks. We therefore propose a detector-based switched model,...
Vulnerable Code Search: Transferable Attack for Code Language Models
Reliable code retrieval is crucial for developer productivity and effective code reuse. However, current neural code language models CLMs powering search tools are susceptible to adversarial attacks targeting non-functional textual elements. In this paper, we introduce a programming...
SSTImap 1.4
SSTImap is a penetration testing software that can check websites for Code Injection and Server-Side Template Injection vulnerabilities and exploit them, giving access to the operating system itself. This tool was developed to be used as an interactive penetration testing tool for SSTI detection...
WPProbe Plugin Enumeration Tool 0.12.11
A fast WordPress plugin and theme scanner that detects installed plugins via REST API enumeration and themes from HTML discovery, then maps them to known vulnerabilities. Over 5,000 plugins detectable without brute-force, thousands more with it...
Answer Is Cheap, Show Me the Evidence! Augmenting Automated Vulnerability Assessment with Evidence
Software vulnerability SV assessment helps prioritize remediation by characterizing reported vulnerabilities. Existing automated methods predict assessment results from SV reports SVRs, but often overlook information in rich text, such as screenshots and code snippets, as well as contextual...
NeuronFuzz: Safety Neuron Guided Fuzzing for LLM Safety Evaluation
Safety evaluation is critical for assessing whether aligned Large Language Models LLMs remain robust against jailbreak attacks. Existing automated testing methods, however, largely rely on response-level feedback: each candidate prompt typically requires generating a target-model response to...
From Fleet to Lab: Revisiting the Security and Complexity of Industrial Rowhammer Mitigation
This paper studies efficient and secure Rowhammer mitigation at the Memory-Controller MC. Rowhammer mitigation faces a fundamental tradeoff between tracking storage and mitigation rate: precise trackers such as Misra-Gries avoid unnecessary mitigations but require large CAM structures, whereas...
PRISM: Lightweight Enclave Isolation with Prismatic Capabilities
Trusted execution environments TEEs protect sensitive code and data from external interference, but lack inherent memory safety. CHERI can enforce spatial memory safety at the object level. Attempts to establish a TEE using CHERI primitives suffer from 1 expensive capability revocation operations...
Security-Aware Pinching-Antenna Systems (PASS): Physical-Layer Security Transmission
This paper investigates heterogeneous secure multi-user transmission in pinching-antenna systems PASS, where dynamically adjustable pinching antennas reshape both guided-wave and free-space propagation to improve communication and confidentiality performance. Unlike conventional physical-layer...
OpenSSL Toolkit 3.6.4
OpenSSL is a robust, fully featured Open Source toolkit implementing the Secure Sockets Layer and Transport Layer Security protocols with full-strength cryptography world-wide. This is the 3.6 release...
FreeBSD Security Advisory - FreeBSD-SA-26:61.openssl
FreeBSD Security Advisory - Multiple issues have been reported as part of this advisory with different issues affecting different OpenSSL versions and therefore different FreeBSD versions...
OpenSSL Toolkit 3.5.8
OpenSSL is a robust, fully featured Open Source toolkit implementing the Secure Sockets Layer and Transport Layer Security protocols with full-strength cryptography world-wide. This is the 3.5 LTS release...
FreeBSD Security Advisory - FreeBSD-SA-26:56.hwpmc
FreeBSD Security Advisory - When a process calls execve2 to execute a setuid or setgid image, hwpmc4 is supposed to detach PMCs owned by unprivileged processes. An inverted check meant that this scenario was not handled properly...
Security Education in Higher Education through AI-Powered Gamification
Cybersecurity education is facing more challenges as AI-driven attacks are becoming increasingly realistic and difficult to detect. Traditional video-based cybersecurity training in higher education often suffers from both low engagement and limited effectiveness. This dilemma motivates educators...
A Drop-In KEM Replacement for Client Signatures in Post-Quantum SSH
The transition to post-quantum cryptography is reshaping the Secure Shell SSH protocol for remote administration. Post-quantum key exchange has been deployed in OpenSSH and is being standardized, while SSH authentication largely remains a signature-replacement effort. This path preserves the...
FreeBSD Security Advisory - FreeBSD-SA-26:58.sound
FreeBSD Security Advisory - The implementation of this ioctl attempts to acquire locks on all channels in a sync group. If locking a channel would block, it releases the sync group list lock and sleeps. Upon reawakening, it is possible that the sync group structure is freed, but the implementatio...
Certified Randomness without Structure against Shallow-Query Adversaries
In a recent breakthrough, Yamakawa and Zhandry J. ACM 2024 constructed a proof of quantumness in the quantum random oracle model QROM in which the quantum prover samples a codeword preimage of a publicly computable function H. They conjectured that given any H, a successful prover must sample the...
Evaluating and Preventing Security Smells in AI-Generated Ansible Code
AI coding assistants generate Infrastructure as Code, yet no work has examined whether this code meets security requirements. This matters because security smells in infrastructure code propagate to deployed systems, producing infrastructure that is insecure and untrustworthy. We evaluate 16 AI...
"Am I Just That Dumb?": Applicability, Action and Verification in Consumer IoT Security Advice
Public campaigns urge people to change default passwords on Internet of Things IoT devices and keep them updated, assuming users can independently determine whether the advice applies. We gave 28 participants in the Netherlands two pieces of government-issued advice reflecting guidance in several...
FreeBSD Security Advisory - FreeBSD-SA-26:60.ppp
FreeBSD Security Advisory - ppp8 contained three memory safety errors in its handling of multilink endpoint discriminator options. mpEnddisc used incorrect length calculations when formatting endpoint discriminator addresses for display, allowing a received endpoint option to overflow a global...
OpenSSL Toolkit 3.4.7
OpenSSL is a robust, fully featured Open Source toolkit implementing the Secure Sockets Layer and Transport Layer Security protocols with full-strength cryptography world-wide. This is the 3.4 release...
FreeBSD Security Advisory - FreeBSD-SA-26:62.tty
FreeBSD Security Advisory - The TIOCSCTTY ioctl handler drops the tty lock in order to acquire the process tree lock. After reacquiring the tty lock, the handler did not revalidate the state of the terminal, and could proceed to link a terminal that was concurrently being destroyed to the calling...
Defending Network Intrusion Detection Systems Based on Graph Neural Networks against Structural Adversarial Attacks
Graph Neural Networks GNNs represent a promising solution for Machine Learning ML based Network Intrusion Detection Systems NIDS, thanks to their ability to leverage both network flow features and topological patterns. While GNN classifiers demonstrate superior robustness against feature-based...
Research Methodologies for Cybersecurity in Enterprise Environments: A Narrative Review, Synthesis and Executable Guide
Enterprise cybersecurity research draws on a wider range of methods than any single community routinely teaches. Researchers face a selection problem before they face a technical one: a study may simultaneously need a systematic review, a design-science artifact, a controlled detection experiment...
Towards LLM-Enhanced Android Taint Analysis
Taint analysis is a fundamental technique for detecting sensitive data leaks in Android apps. However, traditional static tools, such as FlowDroid, still face well-known challenges due to the complexity of accurately modeling the Android framework. In this paper, we investigate whether...
Anatomy of a Scam Call: What 10,000 Real Scam and Spam Calls Reveal about How Phone Scammers Operate
Telephone fraud is pervasive and costly, but its inner workings are rarely observed at scale. We analyze a complete corpus of 10,211 inbound scam and spam calls -- 913 hours of audio and 330,956 transcribed turns from 5,780 distinct numbers -- collected over 54 days by an AI voice-agent honeypot...
BGPay: An Incentive-Compatible Mechanism for BGP Hijack Filtering
BGP hijacking remains a persistent threat as existing defenses, including RPKI/ROV suffer from a fundamental incentive misalignment: the networks best positioned to filter malicious announcements bear operational costs but receive no direct benefit, while the victim prefix owner captures all the...
Joern 4.0.611
Joern is the bug hunter's workbench. With this tool, you can uncover attack surface, sloppy coding practices, and variants of known vulnerabilities using an interactive code analysis shell. Joern supports C, C++, LLVM bitcode, x86 binaries via Ghidra, JVM bytecode via Soot, and Javascript...
Who Falls for SMiSh? Learning through Survey Data Where to Best Target Awareness Training for Mobile Messaging Attacks
As mobile phone adoption has surged, so have scams involving these devices. One such scam, known as SMiShing or smishing after Short Message Service SMS, involves fraudsters sending phishing links via mobile texts. Despite the prevalence of SMiShing, there is a lack of data on who is most...
FreeBSD Security Advisory - FreeBSD-SA-26:57.unix
FreeBSD Security Advisory - The SOCKSTREAM receive path in the unix socket implementation failed to fully detach control messages from the socket buffer before processing them. Some error paths would free those messages, leaving freed data mbufs in the receive socket buffer...
OpenSSL Security Advisory 20260825
OpenSSL Security Advisory 20260825 - ChaCha20-Poly1305 and AES-OCB decryption with an empty ciphertext can report success without verifying the supplied authentication tag when the operation is finalized by calling the EVPCipher function...