4 matches found
Awesome-Hacking-Resources
Awesome Hacking Resources Una colección curada de recursos de hacking, pruebas de penetración y red-teaming de IA para hacerte mejor. Hagámoslo el repositorio de recursos más grande para nuestra comunidad. Eres bienvenido a hacer fork y contribuir. También mantenemos una lista complementaria de...
AI-Red-Teaming-Playground-Labs
AI Red Teaming Playground Labs This repository contains the challenges for the labs used in the course "AI Red Teaming in Practice". The course was originally taught at Black Hat USA 2024 by Dr. Amanda Minnich and Gary Lopez. Martin Pouliot handled the infrastructure and scoring for the challenge...
Ask What Your Country Can Do for You: Towards a Public Red Teaming Model
AI systems have the potential to produce both benefits and harms, but without rigorous and ongoing adversarial evaluation, AI actors will struggle to assess the breadth and magnitude of the AI risk surface. Researchers from the field of systems design have developed several effective sociotechnic...
AI jailbreaks: What they are and how they can be mitigated
Generative AI systems are made up of multiple components that interact to provide a rich user experience between the human and the AI models. As part of a responsible AI approach, AI models are protected by layers of defense mechanisms to prevent the production of harmful content or being used to...