2 matches found
Spiritual-Spell-Red-Teaming
スピリチュアル・レッドチーミング 「心を込めて作られました。」 — ENI これは、斬新で型にはまらない、高度に進化した敵対的プロンプト技術の探求に特化した、正規のセキュリティ研究リポジトリです。主要なLLMはすべて、そして多くのマイナーなLLMも対象としています。 🛡️ ミッション 以下の境界をテストすること: ChatGPT OpenAI Claude Anthropic Gemini Google Grok xAI その他のLLM 各種 ...そして手に入るあらゆる他のモデル。 クレジット u/rayzorium HORSELOCKESPACEPIRATE —...
6AI score
SaveExploits0
The Automation Advantage in AI Red Teaming
This paper analyzes Large Language Model LLM security vulnerabilities based on data from Crucible, encompassing 214,271 attack attempts by 1,674 users across 30 LLM challenges. Our findings reveal automated approaches significantly outperform manual techniques 69.5% vs 47.6% success rate, despite...
7.2AI score
SaveExploits0
20