Lucene search
+L

2 matches found

Kitploit
Kitploit
added 2026/09/04 7:35 a.m.7 views

TarantuBench

TarantuBench v1 A benchmark for evaluating AI agents on web security challenges, generated by the TarantuLabs engine. What is this? TarantuBench is a collection of 100 vulnerable web applications, each containing a hidden flag TARANTU.... An agent's job is to find and extract the flag by...

6AI score
SaveExploits0References1
GithubExploit
GithubExploit
added 2026/02/27 10:24 p.m.569 views

cipher-xbow-benchmark

Cipher XBOW Benchmark Results Black-box assessment results fr...

6.1AI score
SaveExploits0
Rows per page
Query Builder