Lucene search
+L

2 matches found

Packet Storm News
Packet Storm News
•added 2025/06/19 12:00 a.m.•21 views

Probing the Robustness of Large Language Models Safety to Latent Perturbations

Safety alignment is a key requirement for building reliable Artificial General Intelligence. Despite significant advances in safety alignment, we observe that minor latent shifts can still trigger unsafe responses in aligned models. We argue that this stems from the shallow nature of existing...

6.9AI score
SaveExploits0
Packet Storm News
Packet Storm News
•added 2025/05/26 12:00 a.m.•17 views

Novel Loss-Enhanced Universal Adversarial Patches for Sustainable Speaker Privacy

Deep learning voice models are commonly used nowadays, but the safety processing of personal data, such as human identity and speech content, remains suspicious. To prevent malicious user identification, speaker anonymization methods were proposed. Current methods, particularly based on universal...

7.3AI score
SaveExploits0
Rows per page
Query Builder