1 matches found
HelpBench: Assessing the Ability of LLMs to Provide Privacy, Safety, and Security Advice
This paper introduces HelpBench, a benchmark for assessing whether LLMs are capable of providing accurate help in response to questions about digital privacy, safety, and security. We curated 450 questions representing authentic user situations and developed rubrics for each question to evaluate...
5.9AI score
SaveExploits0
20