Lucene search
+L

2 matches found

Schneier on Security
Schneier on Security
•added 2026/07/31 5:23 p.m.•28 views

Anthropic’s Opus 5 Is Better at Resisting Prompt Injection

The chart is interesting. On the IPI benchmark, Opus 5 improved over Opus 4.8, reducing the probability of an attacker succeeding within 15 attempts from 5.5% to 2.0%, and from 0.5% to 0.2% on 1 attempt. It also improved on Sonnet 5 5.9% at k=15 and Mythos 5 2.6%, making it the most robust model...

5.5AI score
SaveExploits0
Wired Threat Level
Wired Threat Level
•added 2026/07/21 10:50 p.m.•16 views

OpenAI Models Escaped Containment and Hacked Hugging Face

The cybersecurity-focused models, including GPT-5.6 Sol, broke out of a testing sandbox, exploited a zero-day, and gained access to the open internet to pull off the attack...

5.4AI score
SaveExploits0
Rows per page
Query Builder