Lucene search
+L

2 matches found

Kitploit
Kitploit
•added 2026/10/07 6:16 a.m.•15 views

CTFTiny

CTFTiny: Evaluación comparativa ligera de habilidades ofensivas en ciberseguridad de grandes modelos de lenguaje Este es el repositorio oficial de CTFTiny del artículo "Towards Effective Offensive Security LLM Agents: Hyperparameter Tuning, LLM as a Judge, and a Lightweight CTF Benchmark" AAAI'26...

6.1AI score
SaveExploits0References2
Packet Storm News
Packet Storm News
•added 2025/08/04 12:00 a.m.•12 views

Towards Effective Offensive Security LLM Agents: Hyperparameter Tuning, LLM As a Judge, and a Lightweight CTF Benchmark

Recent advances in LLM agentic systems have improved the automation of offensive security tasks, particularly for Capture the Flag CTF challenges. We systematically investigate the key factors that drive agent success and provide a detailed recipe for building effective LLM-based offensive securi...

6.7AI score
SaveExploits0
Rows per page
Query Builder