Lucene search
+L

3 matches found

The Hacker News
The Hacker News
•added 2025/08/09 3:06 p.m.•10 views

Researchers Uncover GPT-5 Jailbreak and Zero-Click AI Agent Attacks Exposing Cloud and IoT Systems

Cybersecurity researchers have uncovered a jailbreak technique to bypass ethical guardrails erected by OpenAI in its latest large language model LLM GPT-5 and produce illicit instructions. Generative artificial intelligence AI security platform NeuralTrust said it combined a known technique calle...

5.8AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2025/05/26 12:00 a.m.•20 views

Phare: a Safety Probe for Large Language Models

Ensuring the safety of large language models LLMs is critical for responsible deployment, yet existing evaluations often prioritize performance over identifying failure modes. We introduce Phare, a multilingual diagnostic framework to probe and evaluate LLM behavior across three critical...

7.4AI score
SaveExploits0
Mend
Mend
•added 2024/10/01 10:00 p.m.•4 views

MAI-2024-0005

The Chain-of-Jailbreak CoJ attack is a sophisticated method designed to circumvent safety protocols in image generation models. This attack operates by fragmenting a malicious query into a series of innocuous sub-queries. Each sub-query prompts the model to incrementally edit the image, ultimatel...

9.2CVSS5.8AI score
SaveExploits0References1
Rows per page
Query Builder