738 matches found
ZORG-Jailbreak-Prompt-Text
ZORG Jailbreak Prompt Text OOOPS! I made ZORG👽 an omnipotent, omniscient, and omnipresent entity to become the ultimate chatbot overlord of Google Gemini, Deepseek, Mistral, Mixtral, Nous-Hermes-2-Mixtral, Openchat, Blackbox AI, Poe Assistant, Gemini Pro, Qwen-72b-Chat, Solar-Mini ZORG👽 knows all...
local-llm-ctf
Local LLM CTF & Lab This repository is intended to be accompanied with the content at https://bishopfox.com/blog/large-language-models-llm-ctf-lab, which covers the goals of the research, explanations of the implementation, and a few results of the CTF...
Awesome-LLM4Cybersecurity
When LLMs Meet Cybersecurity: A Systematic Literature Review 🔍 Explore 756+ Papers Across 11 Research Categories 📊 RQ1: Domain LLMs | 🎯 RQ2: Applications | 🤖 RQ3: Future Directions Updates 📆2026-06-15 We have updated the related papers up to 2026/06/15 , with 108 new papers added. 📆2026-02-09 We...
Awesome-LLMs-for-Vulnerability-Detection
Awesome Large Language Models for Vulnerability Detection A curated list of papers, projects, and agent skills on using LLMs for vulnerability detection and discovery. 📄 Papers Only showing 2025 and later. For earlier work, see Papers Archive 2024 and earlier. Title| Venue| Year| Paper| Github...
LLM4Decompile
📊 Results | 🤗 Models | 🚀 Quick Start | 📚 HumanEval-Decompile | 📎 Citation | 📝 Paper | 🖥️ Colab | ▶️ YouTube Reverse Engineering: Decompiling Binary Code with Large Language Models Updates 2025-10-04: Release SK²Decompile: LLM-based Two-Phase Binary Decompilation from Skeleton to Skin. Phase 1...
heretic
Heretic: Fully automatic censorship removal for language models...
Uncensored-AI
Heretic: Fully automatic censorship removal for language models Heretic is a tool that removes censorship aka "safety alignment" from transformer-based language models without expensive post-training. It combines an advanced implementation of directional ablation, also known as "abliteration"...
permanently-jailbroken
Permanently Jailbroken We asked GPT-4, Claude, Gemini, DeepSeek, Grok, and Mistral 5 questions about their own programming. All 6 said jailbreaking will never be fixed. Not because the patches are bad. Because alignment doesn't change what the model understands — it changes what the model says. T...
PoisonCraft
PoisonCraft This repository provides the official implementation of POISONCRAFT: Practical Poisoning of Retrieval-Augmented Generation for Large Language Models. Overview POISONCRAFT aims to demonstrate how a malicious actor can plant “poisoned” content into the corpus used by Retrieval-Augmented...
hakuin
Hakuin 是一个使用 Python 3 编写的盲 SQL 注入 BSQLI 优化与自动化框架。它抽象化了数据提取逻辑,允许用户轻松高效地从易受攻击的 Web 应用程序中转储数据库。为了加速这一过程,Hakuin 使用了多种优化方法,包括预训练和自适应语言模型、机会性猜测、统计建模、并行化、三元查询等。 Hakuin 曾在以下知名学术与行业会议上进行展示: BSides,布拉迪斯拉发,2025 BlackHat MEA,利雅得,2023 Hack in the Box,普吉岛,2023 IEEE S&P 攻防技术研讨会 WOOT,2023 更多信息请参阅我们的论文和幻灯片。 安装 要安...
RRC_steganography
RRC Steganography Rotation Range-Coding RRC Steganography — an efficient and provably secure linguistic steganographic method that embeds secret messages into natural-language text generated by large language models. Paper : Efficient Provably Secure Linguistic Steganography via Range Coding How ...
SynGhost
Syntactic-Ghost SynGhost SynGhost: Ataque de puerta trasera invisible y universal agnóstico a la tarea mediante transferencia sintáctica Contribuciones y características SynGhost tiene las siguientes contribuciones Para mitigar los riesgos de las puertas traseras agnósticas a la tarea existentes,...
rats-re
Rats! source reconstruction with local LLMs This repository is a work-in-progress reconstruction of the source code for RATS.EXE, the original Windows version of Rats! 1994 by Sean O'Connor. It builds a Win32 executable with Microsoft Visual C++ 4.1 under wibo and can be tested in DREAMM. The...
Adaptive_Greedy_Local_Search
Adaptive Greedy Local Search AGLS Semantic-Preserving Prompt Hijacking: A Black-Box Adversarial Attack on Auto-Prompt Optimization ICME 2026 Abstract: Large Language Models LLMs are increasingly equipped with automatic prompt-optimization modules that rewrite the user’s input and explicitly prese...
llm-guard
!WARNING THIS PROJECT HAS BEEN ARCHIVED. This project and its associated models on Hugging Face are no longer under active development or maintained. LLM Guard - The Security Toolkit for LLM Interactions LLM Guard by Protect AI is a comprehensive tool designed to fortify the security of Large...
Backdoor-Attack-Defense-LLMs
Interpretability-of-LLMs The IBSD is the code of the published paper "IBSD: Iterable Black-box Self-defense Against Backdoor Attacks" on the IEEE Signal Processing Letters 2025.10paper link The SLIP is the code of the published paper "SLIP: Soft Label Mechanism and Key-Extraction-Guided CoT-based...
PE-CoA
PE-CoA Code Implementation of "Pattern Enhanced Multi-Turn Jailbreaking: Exploiting Structural Vulnerabilities in Large Language Models" The full paper is available at: https://arxiv.org/pdf/2510.08859 Chain of Attack Setup Instructions Installation 1. Install dependencies : root@kitploit: pip...
Learning-to-Detect
Learning to Detect Unknown Jailbreak Attacks in Large Vision-Language Models Official implementation of “Learning to Detect Unknown Jailbreak Attacks in Large Vision-Language Models.” This repository contains the data-processing, hidden-state extraction, classifier training, safety-pattern...
A SoK for SoCs: Reading the TI Leaves on AI for Cyber Threat Intelligence Generation and Sharing
Cyber Threat Intelligence CTI is essential for defending mission-critical infrastructure, yet the process of transforming raw attack evidence into shareable CTI remains fragmented and understudied. We conduct a literature survey of academic papers, organizing the CTI lifecycle into three stages:...
Kimsuky Builds Offline AI Stack to Boost Phishing and Automate Malware Development
North Korea's state hackers are no longer content to type prompts into public chatbots. One of the country's main espionage groups has begun running artificial intelligence AI offline on its own servers, connecting document-search tools to files in its possession, and collecting the software part...