4 matches found
Awesome-Hacking-Resources
Awesome Hacking Resources A curated collection of hacking, penetration testing, and AI red-teaming resources to make you better. Let's make it the biggest resource repository for our community. You are welcome to fork and contribute. We also maintain a companion tools list — contributions welcome...
AI-Red-Teaming-Playground-Labs
AI Red Teaming Playground Labs このリポジトリには、コース「AI Red Teaming in Practice」で使用されるラボのチャレンジが含まれています。このコースは元々、Black Hat USA 2024 で Dr. Amanda Minnich と Gary Lopez によって教えられました。Martin Pouliot がチャレンジのインフラストラクチャとスコアリングを担当しました。チャレンジは Dr. Amanda Minnich、Gary Lopez、Martin Pouliot...
Ask What Your Country Can Do for You: Towards a Public Red Teaming Model
AI systems have the potential to produce both benefits and harms, but without rigorous and ongoing adversarial evaluation, AI actors will struggle to assess the breadth and magnitude of the AI risk surface. Researchers from the field of systems design have developed several effective sociotechnic...
AI jailbreaks: What they are and how they can be mitigated
Generative AI systems are made up of multiple components that interact to provide a rich user experience between the human and the AI models. As part of a responsible AI approach, AI models are protected by layers of defense mechanisms to prevent the production of harmful content or being used to...