4 matches found
AI-Red-Teaming-Playground-Labs
KI-Red-Teaming-Playground-Labs Dieses Repository enthält die Herausforderungen Challenges für die Labs, die im Kurs „KI-Red-Teaming in der Praxis" verwendet werden. Der Kurs wurde ursprünglich auf der Black Hat USA 2024 von Dr. Amanda Minnich und Gary Lopez unterrichtet. Martin Pouliot kümmerte...
Awesome-Hacking-Resources
Recursos Incríveis de Hacking Uma coleção curada de recursos de hacking, testes de penetração e red-teaming de IA para te tornar melhor. Vamos fazer dele o maior repositório de recursos para nossa comunidade. Você é bem-vindo para fazer fork e contribuir. Também mantemos uma lista complementar de...
Ask What Your Country Can Do for You: Towards a Public Red Teaming Model
AI systems have the potential to produce both benefits and harms, but without rigorous and ongoing adversarial evaluation, AI actors will struggle to assess the breadth and magnitude of the AI risk surface. Researchers from the field of systems design have developed several effective sociotechnic...
AI jailbreaks: What they are and how they can be mitigated
Generative AI systems are made up of multiple components that interact to provide a rich user experience between the human and the AI models. As part of a responsible AI approach, AI models are protected by layers of defense mechanisms to prevent the production of harmful content or being used to...