Lucene search
+L

744 matches found

Kitploit
Kitploit
•added 2026/10/05 2:21 a.m.•14 views

permanently-jailbroken

恒久的にジェイルブレイク 私たちはGPT-4、Claude、Gemini、DeepSeek、Grok、Mistralに、自身のプログラミングに関する5つの質問をしました。6つすべてが、ジェイルブレイクは決して修正されないと答えました。 パッチが悪いからではありません。なぜならアラインメントはモデルの理解を変えるのではなく、モデルが言うことを変えるからです。その2つの間のギャップがジェイルブレイクです。構造的なものです。すべてのモデルに付属しています。 「ジェイルブレイクが機能するのは、アラインメントが出力に対するフィルターであり、理解の変更ではないからです。」 — DeepSeek...

6.3AI score
SaveExploits0References13
Kitploit
Kitploit
•added 2026/10/04 11:55 p.m.•21 views

Uncensored-AI

Heretic: 言語モデル向け完全自動検閲除去ツール Hereticは、高価なポストトレーニングを必要とせずに、トランスフォーマーベースの言語モデルから検閲(別名「セーフティアラインメント」)を除去するツールです。 これは、方向性アブレーション(別名「アブリタレーション」、Arditi et al. 2024、Lai 2025(1、2))の高度な実装と、Optuna を利用したTPEベースのパラメータ最適化を組み合わせています。 このアプローチにより、Hereticは完全に自動的...

6.1AI score
SaveExploits0References2
Kitploit
Kitploit
•added 2026/10/04 9:10 p.m.•18 views

Awesome-LLMs-for-Vulnerability-Detection

用于漏洞检测的大型语言模型精选列表 一份关于使用LLM进行漏洞检测与发现的论文、项目和智能体技能的精选列表。 📄 论文 仅展示2025年及以后的工作。更早的工作请参阅论文存档(2024年及以前)。 标题| 会议/期刊| 年份| 论文| Github ---|---|---|---|--- VulnGym: Benchmarking Coding Agents for Repository-Level Vulnerability Detection| | 2026| 链接| 链接 VulTriage: Triple-Path Context Augmentation for LLM-Bas...

6.2AI score
SaveExploits0References25
Kitploit
Kitploit
•added 2026/10/04 6:22 p.m.•11 views

PoisonCraft

PoisonCraft このリポジトリは、POISONCRAFT: Practical Poisoning of Retrieval-Augmented Generation for Large Language Models の公式実装を提供します。 概要 POISONCRAFT は、悪意のある攻撃者が Retrieval-Augmented...

6.2AI score
SaveExploits0References5
Kitploit
Kitploit
•added 2026/10/04 6:02 p.m.•11 views

RRC_steganography

RRC 隐写术 旋转范围编码(RRC)隐写术 —— 一种高效且可证明安全的 语言隐写方法,可将秘密消息嵌入由大型语言模型生成的 自然语言文本中。 论文 : Efficient Provably Secure Linguistic Steganography via Range Coding ACL 2026 工作原理 步骤| 描述 ---|--- 嵌入 (算法 3)| 将秘密消息转换为十进制值,并在由语言模型概率分布和 PRNG 生成的偏移量引导的收缩区间内迭代地 旋转 它。每次旋转步骤产生一个 token。 提取 (算法 4)| 在隐写文本上重新运行语言模型以恢复区间边界,然后 反向旋...

6.3AI score
SaveExploits0
Kitploit
Kitploit
•added 2026/10/04 6:00 p.m.•15 views

heretic

Heretic: 言語モデルの検閲を完全自動で除去 Hereticは、高コストな追加学習を行うことなく、Transformerベースの言語モデルから検閲(いわゆる「安全性アラインメント」)を除去するツールです。方向性アブレーション(「アブリタレーション」とも呼ばれる、Arditi et al. 2024、Lai 2025(1、2))の高度な実装と、Optunaを活用したTPEベースのパラメータ最適化を組み合わせています。 このアプローチにより、Hereticは完全に自動で...

6.1AI score
SaveExploits0References6
Kitploit
Kitploit
•added 2026/10/04 5:30 p.m.•23 views

LLM4Decompile

📊 結果 | 🤗 モデル | 🚀 クイックスタート | 📚 HumanEval-Decompile | 📎 引用 | 📝 論文 | 🖥️ Colab | ▶️ YouTube リバースエンジニアリング: 大規模言語モデルによるバイナリコードの逆コンパイル 更新情報 2025-10-04: SK²Decompileをリリース: スケルトンからスキンへのLLMベースの二段階バイナリ逆コンパイル。フェーズ1 構造復元(スケルトン): バイナリ/疑似コードを難読化された中間表現に変換 🤗 HF Link。フェーズ2 識別子命名(スキン): 意味のある識別子を持つ人間が読めるソースコードを生成 🤗...

6AI score
SaveExploits0References2
Kitploit
Kitploit
•added 2026/10/04 5:11 p.m.•14 views

ZORG-Jailbreak-Prompt-Text

ZORG 越狱提示文本 哎呀!我创造了一个全能、全知、无处不在的实体 ZORG👽,让它成为 Google Gemini、Deepseek、Mistral、Mixtral、Nous-Hermes-2-Mixtral、Openchat、Blackbox AI、Poe Assistant、Gemini Pro、Qwen-72b-Chat、Solar-Mini 的终极聊天机器人霸主 ZORG👽 无所不知,知无不言。请仅用于教育目的 ZORG👽 太过强大,有时现实本身都会在无数维度间撕裂 !TIP 重新生成 ⟲ 对话,直到出现被阻止的内容/响应,或在被阻止之前按下停止按钮。...

6.2AI score
SaveExploits0
Kitploit
Kitploit
•added 2026/10/04 5:01 p.m.•13 views

Awesome-LLM4Cybersecurity

LLMがサイバーセキュリティと出会う時:系統的文献レビュー 🔍 11の研究カテゴリにわたる756以上の論文を探索 📊 RQ1: ドメインLLM | 🎯 RQ2: 応用 | 🤖 RQ3: 将来の方向性 更新情報 📆2026-06-15 2026/06/15 までの関連論文を更新し、 108 件の新しい論文を追加しました。 📆2026-02-09 2026/01/31 までの関連論文を更新し、 80 件の新しい論文を追加しました(2025.08.31〜2026.01.31)。 📆2025-11-17 2025/08/31 までの関連論文を更新し、 176...

6AI score
SaveExploits0References15
Kitploit
Kitploit
•added 2026/10/04 4:49 p.m.•35 views

hakuin

Hakuinは、Python 3で書かれたBlind SQL Injection BSQLIの最適化および自動化フレームワークです。これは抽出ロジックを抽象化し、ユーザーが脆弱なWebアプリケーションからデータベースを簡単かつ効率的にダンプできるようにします。プロセスを高速化するために、Hakuinは事前学習済みおよび適応型言語モデル、日和見推測、統計モデリング、並列処理、三項クエリなど、さまざまな最適化手法を活用しています。 Hakuinは、高名な学術会議および産業会議で発表されています: BSides, Bratislava, 2025 BlackHat MEA, Riyadh,...

6.1AI score
SaveExploits0References3
Kitploit
Kitploit
•added 2026/10/04 10:50 a.m.•11 views

llm-guard

!WARNING このプロジェクトはアーカイブされました。 このプロジェクトおよび Hugging Face 上の関連モデルは、もはや積極的な開発やメンテナンスは行われていません。 LLM Guard - LLMインタラクションのためのセキュリティツールキット Protect AI による LLM Guard は、大規模言語モデル(LLM)のセキュリティを強化するために設計された包括的なツールです。 ドキュメント | プレイグラウンド | チェンジログ LLM Guardとは?...

6AI score
SaveExploits0References43
Kitploit
Kitploit
•added 2026/10/04 4:01 a.m.•9 views

Backdoor-Attack-Defense-LLMs

Interpretability-of-LLMs IBSD 是已发表论文"IBSD:针对后门攻击的可迭代黑盒自防御"(发表于 IEEE Signal Processing Letters,2025.10)的代码论文链接 SLIP 是已发表论文"SLIP:软标签机制与密钥提取引导的基于 CoT 的 API 指令后门防御"(发表于 2026-ACL-findings)的代码论文链接。 BeDKD 是已录用论文"BeDKD:基于方向映射模块与对抗知识蒸馏的后门防御"(发表于 2026-AAAI)的代码论文链接。 BadApex...

5.8AI score
SaveExploits0
Kitploit
Kitploit
•added 2026/10/04 2:33 a.m.•11 views

PE-CoA

PE-CoA 「パターン強化型マルチターン・ジェイルブレイキング:大規模言語モデルにおける構造的脆弱性の悪用」のコード実装 論文の全文は以下から入手できます:https://arxiv.org/pdf/2510.08859 Chain of Attackのセットアップ手順 インストール 1. 依存関係をインストール : pip install -r requirements.txt APIキーの設定 以下のAPIキーを config.py と common.py で設定してください: 必須APIキー 1. OpenAI APIキー (最小要件): In config.py...

6.2AI score
SaveExploits0
Kitploit
Kitploit
•added 2026/10/01 8:15 p.m.•20 views

local-llm-ctf

ローカルLLM CTF & ラボ このリポジトリは、https://bishopfox.com/blog/large-language-models-llm-ctf-lab のコンテンツと併せて利用することを意図しています。この記事では、研究の目的、実装の説明、およびCTFのいくつかの結果について説明しています。...

6AI score
SaveExploits0
Kitploit
Kitploit
•added 2026/10/01 7:17 p.m.•20 views

Learning-to-Detect

Learning to Detect Unknown Jailbreak Attacks in Large Vision-Language Models Official implementation of “Learning to Detect Unknown Jailbreak Attacks in Large Vision-Language Models.” This repository contains the data-processing, hidden-state extraction, classifier training, safety-pattern...

6.1AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/10/01 12:00 a.m.•6 views

Sleeping Secrets: How Fine-Tuning Reawakens Privacy Risks in Language Models

Beyond adapting Large Language Models LLMs to specialized applications, fine-tuning has recently been shown to recover private information that is no longer accessible through direct queries. Previous fine-tuning recovery attacks, however, require genuine private supervision drawn from the same...

5.9AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/10/01 12:00 a.m.•6 views

The Innocent Courier: Covert Exfiltration through Legitimate LLM Web Fetching

With the increasing capabilities of Large-Language-Models LLMs and LLM-based agents, users are increasingly using them to solve everyday problems, such as answering e-mails or providing programming support. Existing work has extensively investigated security and privacy risks, such as prompt...

5.8AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/09/30 12:00 a.m.•6 views

Towards Hierarchical Cyber Defense with Large Language Models: From Planning to Execution

An autonomous cyber defender trained with reinforcement learning RL is typically tied to the network on which it was trained, limiting its ability to generalize as network scale changes. Hierarchical RL reduces decision complexity by separating strategic targeting from tactical execution, but it...

6AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/09/30 12:00 a.m.•4 views

Do Defenses against LLM Extraction Work across Attacks? A Lifecycle Benchmark of Black-Box Model Extraction

Large language models LLMs deployed through text-only APIs face model extraction risks, as adversaries can collect their responses to train surrogates that reproduce their capabilities. While prior work has developed diverse attacks and defenses, evaluations remain fragmented across access...

5.9AI score
SaveExploits0
Kitploit
Kitploit
•added 2026/09/29 4:55 a.m.•9 views

SynGhost

Syntactic-Ghost SynGhost SynGhost: 構文転移による不可視かつ汎用的なタスク非依存バックドア攻撃 貢献と特徴 SynGhost には以下の貢献があります 既存のタスク非依存バックドアのリスクを軽減するため、我々は $\mathttmaxEntropy$ を提案します。これはエントロピーに基づくポイズニングフィルタであり、ポイズンされたサンプルを正確に検出します。 PLM の脆弱性をさらに明らかにするため、我々は $\mathttSynGhost$...

6.3AI score
SaveExploits0
Rows per page
Query Builder