Lucene search
+L

1 matches found

Packet Storm News
Packet Storm News
added 2025/05/05 12:0 a.m.15 views

RepliBench: Evaluating the Autonomous Replication Capabilities of Language Model Agents

Uncontrollable autonomous replication of language model agents poses a critical safety risk. To better understand this risk, we introduce RepliBench, a suite of evaluations designed to measure autonomous replication capabilities. RepliBench is derived from a decomposition of these capabilities...

7.2AI score
SaveExploits0
Rows per page
Query Builder