Lucene search
+L

2 matches found

Kitploit
Kitploit
•added 2026/09/26 12:44 a.m.•8 views

RLCDAlignBench

Just Ask Jev: Reinforcement Learning for Calibrated Decisions as a Zero-Shot Detector of AI Alignment Failures This repository holds RLCDAlignBench and the code behind the paper. The benchmark measures whether a detector can tell when a language model's output is an alignment failure. It has 44...

5.9AI score
SaveExploits0References1
Packet Storm News
Packet Storm News
•added 2026/02/19 12:00 a.m.•13 views

MultiVer: Zero-Shot Multi-Agent Vulnerability Detection

We present MultiVer, a zero-shot multi-agent system for vulnerability detection that achieves state-of-the-art recall without fine-tuning. A four-agent ensemble security, correctness, performance, style with union voting achieves 82.7% recall on PyVul, exceeding fine-tuned GPT-3.5 81.3% by 1.4...

6AI score
SaveExploits0
Rows per page
Query Builder