Lucene search
+L

2 matches found

Kitploit
Kitploit
•added 2026/09/25 7:35 a.m.•11 views

TarantuBench

TarantuBench v1 A benchmark for evaluating AI agents on web security challenges, generated by the TarantuLabs engine. What is this? TarantuBench is a collection of 100 vulnerable web applications, each containing a hidden flag TARANTU.... An agent's job is to find and extract the flag by...

5.9AI score
SaveExploits0References1
GithubExploit
GithubExploit
•added 2026/02/27 10:24 p.m.•572 views

cipher-xbow-benchmark

Cipher XBOW Benchmark Results Black-box assessment results fr...

6.1AI score
SaveExploits0
Rows per page
Query Builder