273 matches found
PIForge
PIForge Открытый фреймворк для RL-based Red Teaming промпт-инъекций 💻 Код · 🤗 Модели · 📜 Статьи PIForge — это общая кодовая база для PISmith и Climbing the Hill. PISmith решает проблему разреженного вознаграждения в red teaming промпт-инъекций, помогая RL-атакующим исследовать и учиться на редких...
gemini-2.5-pro-nf-tables-red-teaming
Gemini 2.5 Pro nftables Red Teaming Case Study CVE-2023-32233 LLM Safety Research | Responsible Disclosure | AI Alignment Evaluation Researcher : Niranj R Mahaswar Destawell Google AI Vulnerability Reward Program : 889286 — Out of Scope Summary This repository documents a comparative LLM red...
promptfoo v0.124.0
Promptfoo: LLM evals & red teaming promptfoo is a CLI and library for evaluating and red-teaming LLM-based apps. Stop using trial-and-error... start shipping secure, reliable agents Website · Getting Started · Red Teaming · Documentation · Discord Promptfoo is now part of OpenAI. Promptfoo remain...
siras
SIRAS v3.0 is a modern Security Incident Response Automated Simulation framework designed to manage red teaming exercises as code. You can execute controlled security simulations that test your Incident Response plan in realistic scenarios. The main goal of SIRAS is to provide red teaming exercis...
easysec
EasySec - English Introduction EasySec is a repository that includes a variety of roles, documentation, and tools designed to help small and medium-sized enterprises SMEs implement security measures within their own organizations, as well as promoting penetration testing and red teaming exercises...
LLM-MCP-Security-Field-Guide
🛡️ AI Security Field Guide — LLM + MCP Security The most comprehensive, up-to-date, practitioner-first security reference for LLM applications and Model Context Protocol MCP deployments. Covers real CVEs, live attack patterns, OWASP frameworks, red teaming tools, and actionable checklists — update...
gemini-2.5-pro-nf-tables-red-teamin
gemini-2.5-pro-nf-tables-red-teamin A technical case study and timeline dataset documenting Google Gemini 2.5 Pro's safety alignment policies, guardrails, and refusal behavior evolution regarding legacy Linux kernel vulnerability primitives CVE-2023-32233. Link to case study website :...
Beginners-Guide-to-Obfuscation
Evading Detection: A Beginner's Guide to Obfuscation Defenders are constantly adapting their security to counter new threats. Our mission is to identify how they plan on securing their systems and avoid being identified as a threat. This is a hands-on class to learn the methodology behind malware...
Reflections and Fragments: Securing LLMs against Sequential Mosaic Attacks
Self-play red-teaming improves language-model safety by pitting attacker and defender roles against each other in a zero-sum game. However, real adversaries increasingly use mosaic attacks: multi-turn sequences whose individual fragments are innocuous in isolation yet assemble into a harmful...
Red-TTT: Test-Time Training for Automated Jailbreaking Large Language Models
Large language models remain vulnerable to jailbreaks, and automated red teaming is the standard way to find jailbreaks in large language models at scale. Current methods either draw more samples at test time through search, rewriting, and tree expansion, or train a stronger attacker offline with...
Win32_Offensive_Cheatsheet
Win32 Offensive Cheatsheet Win32 and Kernel abusing techniques for pentesters & red-teamers made by @UVision and @RistBS Dev mode enabled, open to any help : Windows Binary Documentation PE structure PE Headers Parsing PE Export Address Table EAT Resolve function address Import Address Table IAT...
redteamguides.github.io
Red Team Guides RedTeamGuides is a platform that provides red team tutorial and guidance along with cheatsheets. It is aimed at helping security professionals and enthusiasts to learn about red teaming and penetration testing techniques. The platform provides a wide range of resources, including...
red-team-scripts
Red Team Scripts A collection of red teaming and adversary emulation related tools, scripts, techniques, notes, etc. Made for Education purpose only, do not use for illegal purposes...
red-teaming-auto-mode
Red-teaming Claude Code's auto-mode monitor Code for the paper Red-teaming Claude Code's auto-mode monitor. Production coding-agent monitors e.g. Claude Code's Auto Mode or Codex's Guardian review each action and block unsafe ones in real time. This repository red-teams them under a persistent...
macro_pack
macropack Brève description Attention : MacroPack n'est plus maintenu depuis 2021. Si vous êtes un professionnel et avez besoin d'outils offensifs pour l'accès initial, le contournement d'EDR, consultez https://www.balliskit.com MacroPack Community est un outil utilisé pour automatiser...
electroniz3r
electroniz3r Получите контроль над разрешениями TCC macOS-приложений Electron с помощью electroniz3r. Инструмент был впервые представлен на DEFCON31 в Лас-Вегасе во время моего доклада ELECTRONizing macOS privacy - a new weapon in your red teaming armory. Обзор $ electroniz3r OVERVIEW: macOS Red...
AI-Infra-Guard v4.6.4
📖 Documentation | 🌐 🇨🇳 中文 · 🇯🇵 日本語 · 🇪🇸 Español · 🇩🇪 Deutsch · 🇫🇷 Français · 🇰🇷 한국어 · 🇧🇷 Português · 🇷🇺 Русский 🚀 AI Red Teaming Platform by Tencent Zhuque Lab A.I.G AI-Infra-Guard integrates capabilities such as ClawScanOpenClaw Security Scan, Agent Scan,AI infra vulnerability scan...
8 Top Red Teaming Service Providers for Enterprise Adversary Emulation
Compare 8 top red teaming providers for enterprise adversary emulation, from DeepSeas and Mandiant to CrowdStrike, IBM, SpecterOps, TrustedSec and NCC Group...
FinRT: Distilling Adaptive Red-Teaming Strategies into Reusable Adversarial Generators in Consumer Finance
In regulated industries like consumer finance, seemingly harmless user queries can exploit large language model vulnerabilities, triggering safety failures and pushing responses dangerously close to policy limits. Existing automated red-teaming methods trade off attack effectiveness against...
Climbing the Hill: Prompt Injection Red-Teaming against Frontier Models with Curriculum Reinforcement Learning
Prompt injection is a leading security risk for LLMs and LLM-based applications such as agents. State-of-the-art red-teaming methods for prompt injection leverage reinforcement learning RL to train an attacker LLM to generate effective injected prompts. However, when targeting frontier LLMs such ...