Lucene search
+L

315 matches found

Kitploit
Kitploit
•added 2026/10/10 3:14 p.m.•19 views

Judge-Jury-and-Executable

Judge Jury and Executable एक थ्रेट हंटिंग फोरेंसिक टूल विशेषताएँ: किसी माउंटेड फाइलसिस्टम को तुरंत खतरों के लिए स्कैन करें या किसी घटना से पहले सिस्टम बेसलाइन एकत्र करें, अतिरिक्त थ्रेट हंटिंग क्षमता के लिए किसी घटना से पहले, उसके दौरान या बाद में उपयोग किया जा सकता है एक से लेकर कई वर्कस्टेशनों ...

6.6AI score
SaveExploits0
Kitploit
Kitploit
•added 2026/10/08 9:42 a.m.•10 views

FinRED-paper

FinRED: An Expert-Guided Benchmark Generation and Evaluation Framework for Financial LLM Red-Teaming IEEE ICDM 2026 A red-team benchmark generation pipeline for safety evaluation in the financial domain. Supplementary Documentation Detailed materials referenced in the paper:...

6.2AI score
SaveExploits0References4
Kitploit
Kitploit
•added 2026/10/06 4:55 a.m.•17 views

Rubrics-as-an-Attack-Surface

المعايير كسطح هجومي: انحراف التفضيلات الخفي في حكام النماذج اللغوية الكبيرة 📊 مجموعة البيانات • 🤖 النماذج المدرَّبة • 📝 الورقة البحثية • 💻 المستودع يحتوي هذا المستودع على الكود الخاص بالورقة البحثية المعايير كسطح هجومي: انحراف التفضيلات الخفي في حكام النماذج اللغوية الكبيرة من تأليف Ruomeng Ding،...

6.3AI score
SaveExploits0References1
Packet Storm News
Packet Storm News
•added 2026/09/21 12:00 a.m.•17 views

Attack Success Rate Is Not a Number: On Measurement Validity in Agentic AI Security Evaluation

Attack success rate ASR is the headline metric in nearly every published evaluation of attacks on, and defenses for, LLM agents. We argue that ASR as currently used is not a single quantity but a family of metrics parameterized by six design choices that papers seldom specify and never hold...

5.8AI score
SaveExploits0
Cvelist
Cvelist
•added 2026/09/15 2:45 p.m.•40 views

CVE-2026-55617 Hydro: Insufficient session expiration when recreating sessions

Hydro is a next-generation high-performance online judge platform. From 4.10.4 until 5.0.2, the session recreation logic in packages/hydrooj/src/service/layers/base.ts creates a replacement session token without deleting the previous token from the server-side session token store, so an old sid...

6.9CVSS0.00473EPSS
SaveExploits0References4
GithubExploit
GithubExploit
•added 2026/09/05 6:54 p.m.•23 views

exploitgym-unified

ExploitGym Unified Runner - 898 instâncias + Dashboard Live R...

6AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/09/01 12:00 a.m.•11 views

Ranked by the Matcher: A Reproducibility Audit of Knowledge Graph Extraction from Threat Reports

Security teams and researchers choose knowledge-graph extraction tooling for threat reports on the strength of published triple-F1 scores, yet those scores depend on how predicted triples are matched to gold annotations. We could reimplement the stated matching rule for only five of twelve...

5.9AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/08/04 12:00 a.m.•147 views

CLEAR: Causal Context-Based Agentic Reasoning for Vulnerability Detection

Detecting source code vulnerabilities is increasingly difficult as modern security flaws are rooted in complex causal dependencies between execution flows, control conditions, and program states. Despite recent advances in Large Language Models LLMs and multi-agent frameworks, existing approaches...

5.9AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/07/29 12:00 a.m.•20 views

MemSecBench: Tracking Agent Memory Poisoning from Persistence to Consequence and Repair

Memory systems allow agents to retain and reuse information from past interactions, but they can also let malicious content persist. A malicious instruction crafted by an attacker may be stored in long-term memory, recalled much later, and quietly shape a real action. Recent benchmarks increasing...

5.9AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/07/28 12:00 a.m.•24 views

SkillGate: Cost Efficient Runtime Malicious Skill File Detection in Coding Agents

Software engineering teams now deploy AI coding agents Cursor, Claude Code, GitHub Copilot as first-class productivity tools, installing domain-specific skill files to tailor agent behavior to project APIs, framework conventions, and organizational workflows. These complex Markdown files are easi...

5.9AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/07/28 12:00 a.m.•77 views

StealthBench: Measuring Operational Stealth in Autonomous Offensive-Security Agents

Stealth, the discipline of achieving an objective without revealing your presence, capabilities, or collected intelligence, is what separates sophisticated operators from detectable ones. Elite security researchers and advanced persistent threats achieve their objectives unnoticed; autonomous...

5.9AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/07/25 12:00 a.m.•15 views

From Signals to Behaviors: Evidence-Based Android Malware Detection

Android malware remains a persistent threat, and detecting it accurately is a long-standing open problem. Whether an app is malicious depends on what it actually does and the context in which it does it, not on the surface signals it happens to exhibit. Existing detectors instead reason about...

6AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/07/15 12:00 a.m.•43 views

ProfMalPlus: Agent-Coordinated Detection of Malicious NPM Packages Via Static-Dynamic Analysis Synergy

Open source software is vulnerable to supply-chain attacks through transitive dependencies, especially malicious code injected into NPM packages. Existing detectors often inadequately model obfuscated behavior, overlook JavaScript's object-centric features, poorly coordinate static and dynamic...

6.1AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/07/08 12:00 a.m.•14 views

Beyond Attack-Success Rate: Action-Graded Severity Scale for Tool-Using AI Agents

Agentic red-teaming benchmarks report whether an injected agent was compromised as a single bit: the attack succeeded, or it did not. We argue that this binary attack-success rate discards the information a defender most needs, namely how harmful the resulting action was. We introduce an...

6AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/06/22 12:00 a.m.•27 views

Detecting Malicious Agent Skills in the Wild Using Attention

LLM agents increasingly load skills, file-based packages of natural-language instructions written by third parties and distributed through marketplaces, that execute with the user's privileges. A single malicious skill can exfiltrate data, hijack the agent, or persist as a supply-chain foothold,...

5.9AI score
SaveExploits0
GithubExploit
GithubExploit
•added 2026/06/16 3:37 p.m.•134 views

TrustedRouter-ExploitBench

TrustedRouter-ExploitBench Notes, harness configs, and a runb...

5.5AI score
SaveExploits0
RedhatCVE
RedhatCVE
•added 2026/06/05 7:33 p.m.•25 views

CVE-2026-9528

A vulnerability was identified in itsourcecode Electronic Judging System 1.0. Impacted is an unknown function of the file /admin/deletejudge.php. Such manipulation of the argument judgeid leads to sql injection. The attack can be executed remotely. The exploit is publicly available and might be...

7.5CVSS7.1AI score0.00411EPSS
SaveExploits0References1
Packet Storm News
Packet Storm News
•added 2026/06/03 12:00 a.m.•121 views

ZERO-APT: A Closed-Loop Adversarial Framework for LLM-Driven Automated Penetration Testing under Intelligent Defense

LLM-driven automated penetration testing agents are typically evaluated against static targets that neither detect nor respond to attacks, so their behavior under intelligent defense remains untested. The causal consistency of multi-step attack chains likewise hinges on unstable LLM reasoning, an...

5.5AI score
SaveExploits0
RedhatCVE
RedhatCVE
•added 2026/05/28 8:13 p.m.•24 views

CVE-2026-9525

A vulnerability has been found in itsourcecode Electronic Judging System 1.0. This affects an unknown part of the file /admin/editjudge.php. The manipulation of the argument judgeid leads to sql injection. The attack may be initiated remotely. The exploit has been disclosed to the public and may ...

7.5CVSS6.8AI score0.00411EPSS
SaveExploits0References1
NVD
NVD
•added 2026/05/26 5:16 a.m.•27 views

CVE-2026-9528

A vulnerability was identified in itsourcecode Electronic Judging System 1.0. Impacted is an unknown function of the file /admin/deletejudge.php. Such manipulation of the argument judgeid leads to sql injection. The attack can be executed remotely. The exploit is publicly available and might be...

7.5CVSS0.00411EPSS
SaveExploits0References5
Rows per page
Query Builder