1 matches found
Trust Me, I'M Your Developer: Self-Issued Authentication in Large Language Models
Large language model LLM security has largely focused on role-playing jailbreaks, with less attention to what happens when a user asks an LLM to verify an identity claim through a test designed by the model itself. We study this behavior through a staged developer-identity experiment with ChatGPT...
5.5AI score
SaveExploits0
20