Lucene search
+L

4 matches found

The Hacker News
The Hacker News
added 7 hours ago9 views

Anthropic Says Claude Mistook the Open Internet for a CTF and Breached Three Organizations

Anthropic on Thursday became the latest artificial intelligence AI company to reveal that three of its models, including Claude Opus 4.7, Mythos 5, and an unnamed research model, had breached three unnamed organizations during cybersecurity testing without its knowledge. The AI firm said the...

5.8AI score
SaveExploits0
Packet Storm News
Packet Storm News
added 2026/04/14 12:0 a.m.14 views

Honeypot Protocol

Trusted monitoring, the standard defense in AI control, is vulnerable to adaptive attacks, collusion, and strategic attack selection. All of these exploit the fact that monitoring is passive: it observes model behavior but never probes whether the model would behave differently under different...

5.8AI score
SaveExploits0
The Hacker News
The Hacker News
added 2026/03/07 11:21 a.m.17 views

Anthropic Finds 22 Firefox Vulnerabilities Using Claude Opus 4.6 AI Model

Anthropic on Friday said it discovered 22 new security vulnerabilities in the Firefox web browser as part of a security partnership with Mozilla. Of these, 14 have been classified as high, seven have been classified as moderate, and one has been rated low in severity. The issues were addressed in...

9.8CVSS5.8AI score0.00625EPSS
SaveExploits2
The Hacker News
The Hacker News
added 2026/02/06 5:49 a.m.10 views

Claude Opus 4.6 Finds 500+ High-Severity Flaws Across Major Open-Source Libraries

Artificial intelligence AI company Anthropic revealed that its latest large language model LLM, Claude Opus 4.6, has found more than 500 previously unknown high-severity security flaws in open-source libraries, including Ghostscript, OpenSC, and CGIF. Claude Opus 4.6, which was launched Thursday,...

6.3AI score
SaveExploits0
Rows per page
Query Builder