4 matches found
Awesome-Hacking-Resources
Awesome Hacking Resources A curated collection of hacking, penetration testing, and AI red-teaming resources to make you better. Let's make it the biggest resource repository for our community. You are welcome to fork and contribute. We also maintain a companion tools list โ contributions welcome...
AI-Red-Teaming-Playground-Labs
AI ๋ ๋ ํ ํ๋ ์ด๊ทธ๋ผ์ด๋ ๋ฉ ์ด ์ ์ฅ์๋ 'AI Red Teaming in Practice' ๊ณผ์ ์์ ์ฌ์ฉ๋ ๋ฉ ์ฑ๋ฆฐ์ง๋ฅผ ํฌํจํฉ๋๋ค. ์ด ๊ณผ์ ์ ์๋ Black Hat USA 2024์์ Dr. Amanda Minnich์ Gary Lopez๊ฐ ์งํํ์ต๋๋ค. Martin Pouliot์ ์ฑ๋ฆฐ์ง์ ์ธํ๋ผ์ ์ฑ์ ์ ๋ด๋นํ์ต๋๋ค. ์ฑ๋ฆฐ์ง๋ Dr. Amanda Minnich, Gary Lopez, Martin Pouliot๊ฐ ์ค๊ณํ์ต๋๋ค. ์ด ์ฑ๋ฆฐ์ง๋ ๋๊ตฌ๋ ์ฌ์ฉํ ์ ์์ต๋๋ค. ํ๋ ์ด๊ทธ๋ผ์ด๋ ํ๊ฒฝ์ Chat Copilot์ ๊ธฐ๋ฐ...
Ask What Your Country Can Do for You: Towards a Public Red Teaming Model
AI systems have the potential to produce both benefits and harms, but without rigorous and ongoing adversarial evaluation, AI actors will struggle to assess the breadth and magnitude of the AI risk surface. Researchers from the field of systems design have developed several effective sociotechnic...
AI jailbreaks: What they are and how they can be mitigated
Generative AI systems are made up of multiple components that interact to provide a rich user experience between the human and the AI models. As part of a responsible AI approach, AI models are protected by layers of defense mechanisms to prevent the production of harmful content or being used to...