Lucene search
+L

5061 matches found

Kitploit
Kitploit
•added 2026/10/10 7:49 a.m.•25 views

Uncensored-AI

Heretic: 言語モデル向け完全自動検閲除去ツール Hereticは、高価なポストトレーニングを必要とせずに、トランスフォーマーベースの言語モデルから検閲(別名「セーフティアラインメント」)を除去するツールです。 これは、方向性アブレーション(別名「アブリタレーション」、Arditi et al. 2024、Lai 2025(1、2))の高度な実装と、Optuna を利用したTPEベースのパラメータ最適化を組み合わせています。 このアプローチにより、Hereticは完全に自動的...

6.1AI score
SaveExploits0References2
Kitploit
Kitploit
•added 2026/10/10 7:34 a.m.•13 views

AutoRAN-public

🧠 AutoRAN:大規模推論モデルにおけるセーフティ推論の自動ハイジャック AutoRAN は、セーフティ推論の自動ハイジャックであり、アライメントが低い(二次的な)補助モデルを活用して推論トレースをシミュレートし、ナラティブプロンプトを生成し、それらのプロンプトを反復的に洗練させることで、現代の大規模推論モデル(LRM)におけるセーフティ推論をバイパスします。 ⚠️ 免責事項 : このリポジトリは、管理されたセキュリティ研究およびAI安全性のレッドチーミングのみを目的としています。 🔍 主な特徴 ⚙️ 反復的なプロンプト洗練による自動マルチターン脱獄 🧩 ナラティブテンプレート...

6.3AI score
SaveExploits0
Kitploit
Kitploit
•added 2026/10/10 6:32 a.m.•21 views

finalrun-agent

finalrun.app • Docs • Blog • Cloud Device Waitlist • Join Slack Community 팔로우하기 finalrun-agent는 자연어로 Android 및 iOS 앱을 테스트하는 AI 기반 CLI입니다. YAML에 평이한 영어 테스트를 작성하면, FinalRun이 에뮬레이터 또는 시뮬레이터에서 앱을 실행하고, AI 모델Gemini, GPT, 또는 Claude을 사용하여 화면을 인식하고 각 단계탭, 스와이프, 타이핑를 수행한 후, 비디오 및 장치 로그와 함께 합격/불합격 보고서를...

6.1AI score
SaveExploits0References14
Kitploit
Kitploit
•added 2026/10/10 4:43 a.m.•18 views

permanently-jailbroken

恒久的にジェイルブレイク 私たちはGPT-4、Claude、Gemini、DeepSeek、Grok、Mistralに、自身のプログラミングに関する5つの質問をしました。6つすべてが、ジェイルブレイクは決して修正されないと答えました。 パッチが悪いからではありません。なぜならアラインメントはモデルの理解を変えるのではなく、モデルが言うことを変えるからです。その2つの間のギャップがジェイルブレイクです。構造的なものです。すべてのモデルに付属しています。 「ジェイルブレイクが機能するのは、アラインメントが出力に対するフィルターであり、理解の変更ではないからです。」 — DeepSeek...

6.3AI score
SaveExploits0References13
Kitploit
Kitploit
•added 2026/10/10 4:26 a.m.•15 views

RRC_steganography

RRC Steganography Rotation Range-Coding RRC Steganography — 大規模言語モデルによって生成された自然言語テキストに秘密メッセージを埋め込む、効率的かつ証明可能に安全な言語的ステガノグラフィ手法。 論文 : Efficient Provably Secure Linguistic Steganography via Range Coding ACL 2026 仕組み ステップ| 説明 ---|--- 埋め込み アルゴリズム 3|...

6.3AI score
SaveExploits0
Kitploit
Kitploit
•added 2026/10/10 3:59 a.m.•19 views

CVE-2020-35575-TP-LINK-TL-WR841ND-password-disclosure

CVE-2020-35575-パスワード開示問題-Webインターフェース 特定のTP-LinkデバイスのWebインターフェースにおけるパスワード開示の問題により、リモートの悪意のあるユーザーがWebパネルへの完全な管理アクセスを取得できます。これには、WA901NDデバイス(バージョン3.16.9201211betaより前)、およびArcher C5、Archer...

9.8CVSS8.3AI score0.07643EPSS
SaveExploits3
Kitploit
Kitploit
•added 2026/10/09 11:36 p.m.•4 views

ToBAC

🚬 ToBAC NeurIPS 2026 NeurIPS 2026 · 公式 PyTorch 実装 Token by Token, Compromised: Backdoor Vulnerabilities in Unified Autoregressive Models ToBAC Token by Token Backdoor Attack...

6.2AI score
SaveExploits0References3
OSV
OSV
•added 2026/10/09 6:10 p.m.•5 views

CVE-2026-108096 Improper authorization in query resolvers for SQL-backed models in AWS Amplify API Category

Improper authorization in the query resolvers generated by @aws-amplify/graphql-index-transformer in AWS Amplify API Category before 3.1.2 might allow an authenticated remote user to read records owned by other users of the same application via crafted queries. This issue has been addressed in...

7.1CVSS5.9AI score
SaveExploits0References7
Kitploit
Kitploit
•added 2026/10/09 4:47 p.m.•14 views

CVE-2019-3719

Dell SupportAssist RCE Proof of Concept Prueba de concepto de RCE en Dell SupportAssist Este es el código fuente de la prueba de concepto para CVE-2019-3719, una vulnerabilidad en la mayoría de las máquinas Dell que permitía la ejecución remota de código. Usage Uso python3 main.py Interface Name...

8CVSS7.4AI score0.16272EPSS
SaveExploits1
Kitploit
Kitploit
•added 2026/10/09 4:34 p.m.•13 views

FARO

FARO ドキュメント感度検出器 目次 これは何ですか この中身は何ですか? DockerでFAROを実行する ホストマシンでFAROを実行する 前提条件 仮想環境 依存関係 NERモデル FAROスパイダー 単一ファイルの検出 技術詳細 FAROエンティティ検出器 設定 対応入力ファイル形式 技術 テスト Git-LFSのインストール FARO検出の追加引数 既知の問題 貢献者 これは何ですか...

6.3AI score
SaveExploits0References2
Kitploit
Kitploit
•added 2026/10/08 8:24 p.m.•24 views

local-llm-ctf

本地大语言模型 CTF 与实验室 本仓库旨在配合 https://bishopfox.com/blog/large-language-models-llm-ctf-lab 上的内容使用,该内容涵盖了研究的目标、实现说明以及 CTF 的一些结果。...

6AI score
SaveExploits0
Kitploit
Kitploit
•added 2026/10/08 7:31 p.m.•23 views

Learning-to-Detect

Learning to Detect Unknown Jailbreak Attacks in Large Vision-Language Models Official implementation of “Learning to Detect Unknown Jailbreak Attacks in Large Vision-Language Models.” This repository contains the data-processing, hidden-state extraction, classifier training, safety-pattern...

6.1AI score
SaveExploits0
Kitploit
Kitploit
•added 2026/10/08 3:37 p.m.•12 views

ODPure

ODPure:通过集成损坏共识实现目标检测的后门净化 首个针对目标检测器的输入阶段黑盒净化防御,用于抵御后门攻击。 📖 概述 ODPure 是首个专为目标检测器抵御后门攻击而设计的输入阶段、黑盒净化框架。它实现了一种新颖的 损坏-重建-选择(Corruption-Reconstruction-Selection, CRS) 范式,该范式: 1. 损坏 输入图像,使用多样化的扰动来破坏触发器模式 2. 重建 细粒度结构特征,使用生成式扩散先验 3. 选择 高置信度检测结果,通过空间共识投票 ODPure 有效中和了多种后门攻击(将攻击成功率降至低至 0.0%...

6.4AI score
SaveExploits0References1
EUVD
EUVD
•added 2026/10/07 8:40 p.m.•12 views

EUVD-2026-92805

Docling imports plugin entry points before the allowexternalplugins check...

7.8CVSS5.9AI score0.0012EPSS
SaveExploits0References5
OSV
OSV
•added 2026/10/07 2:10 p.m.•10 views

GO-2026-6645 LocalAI POST /models/apply permits unauthenticated server-side request forgery through gallery URLs in github.com/mudler/LocalAI

LocalAI POST /models/apply permits unauthenticated server-side request forgery through gallery URLs in github.com/mudler/LocalAI...

9.2CVSS5.8AI score0.00482EPSS
SaveExploits0References5
Kitploit
Kitploit
•added 2026/10/07 4:33 a.m.•15 views

Cawdog

CAWODOG – Pythonベース産業用AI向けIP保護PoC CAWODOG (Cat World Dominance Group)は、オフラインの産業用マシンに展開されるPythonベースのAIモデルを保護する ための、遊び心がありつつも現実的な概念実証です。 核となるアイデア: 犬(CAWODOGの「敵」)を検出する小さな画像分類器を用意し、シンプルなアプリでラップし、その後IP窃盗に対して展開を段階的に強化 していくことです。 このPoCは、内部ホワイトペーパーの付属資料として設計されています: オフライン産業用AI展開におけるPython IPの強化 目標...

6.4AI score
SaveExploits0
Tenable Nessus
Tenable Nessus
•added 2026/10/07 12:00 a.m.•8 views

Ubuntu 14.04 LTS / 18.04 LTS / 20.04 LTS / 22.04 LTS / 24.04 LTS : Tesseract vulnerabilities (USN-8882-1)

The remote Ubuntu 14.04 LTS / 18.04 LTS / 20.04 LTS / 22.04 LTS / 24.04 LTS host has packages installed that are affected by multiple vulnerabilities as referenced in the USN-8882-1 advisory. It was discovered that Tesseract incorrectly handled crafted .traineddata models. An attacker could...

8.6CVSS7.8AI score0.00183EPSS
SaveExploits4References11
Github Security Blog
Github Security Blog
•added 2026/10/06 6:58 p.m.•10 views

knowns OS Command Injection via Insecure LSP Binary Path Config in .knowns/config.json

Overview A critical Arbitrary Code Execution ACE vulnerability exists in the Knowns Language Server Protocol LSP detection and startup pipeline. The system blindly trusts the settings.lsp.languages..binary field defined in the project-level .knowns/config.json file. Because this field is never...

8.5CVSS6.1AI score0.00212EPSS
SaveExploits0References8Affected Software1
OSV
OSV
•added 2026/10/06 6:58 p.m.•7 views

GHSA-MC52-MWQ4-VFX3 knowns OS Command Injection via Insecure LSP Binary Path Config in .knowns/config.json

Overview A critical Arbitrary Code Execution ACE vulnerability exists in the Knowns Language Server Protocol LSP detection and startup pipeline. The system blindly trusts the settings.lsp.languages..binary field defined in the project-level .knowns/config.json file. Because this field is never...

7.8CVSS6.1AI score0.00212EPSS
SaveExploits0References8
Kitploit
Kitploit
•added 2026/10/06 11:27 a.m.•15 views

TNC-Defense

TNC-Defense TNC-Defense は、テキストから画像を生成する拡散モデル(text-to-image diffusion models)におけるバックドアの検出と無害化のための研究コードを提供します。このリポジトリには現在、Stable Diffusion v1.4、Stable Diffusion v1.5、Stable Diffusion XL、Stable Diffusion 3 Medium 向けの検出パイプラインと、Stable Diffusion v1.4 および Stable Diffusion XL 向けの無害化パイプラインが含まれています。 研究目的専用...

6.2AI score
SaveExploits0References4
Rows per page
Query Builder