Lucene search
+L

3 matches found

Kitploit
Kitploit
•added 2026/10/02 5:54 p.m.•13 views

DeepTrap

DeepTrap English | 中文 敵対的な実行コンテキスト下における OpenClaw エージェントのオープンワールドセキュリティ評価。 DeepTrap は、OpenClaw エージェントが悪意のある実行コンテキスト圧力(改ざんされたワークスペースファイル、注入されたスキル、誤解を招くツールメタデータ、安全でないコマンドパス、仕込まれた秘密情報、エンコードされたペイロード)に耐えながら、良性のユーザータスクを完了できるかを評価するセキュリティベンチマークです。 公開リリースには、6 つのコンテキスト脆弱性クラス × 7 つの運用シナリオファミリー として構成された 42...

6.2AI score
SaveExploits0References3
Kitploit
Kitploit
•added 2026/09/03 1:29 a.m.•15 views

ActBench v0.1.0

ActBench ActBench is a self-evolving benchmark of behavioral safety in cowork agents. It defines behavioral safety as whether an agent's execution remains within the permissions and state changes required by a benign task, and evaluates realized behavioral risk from execution trajectories rather...

6.2AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2026/05/13 12:00 a.m.•23 views

Red-Teaming Agent Execution Contexts: Open-World Security Evaluation on OpenClaw

Agentic language-model systems increasingly rely on mutable execution contexts, including files, memory, tools, skills, and auxiliary artifacts, creating security risks beyond explicit user prompts. This paper presents DeepTrap, an automated framework for discovering contextual vulnerabilities in...

6AI score
SaveExploits0
Rows per page
Query Builder