2 matches found
JEV As a Judge for Agent Trace Security: An Empirical Comparison with Generative LLM Judges
Security evaluation of tool-using agents requires judging actions in context, yet generative judges add latency, explanation overhead, and output-validation failures. We study whether JEV, a typed decision model, offers a useful alternative for retrospective trace classification. We evaluate JEV...
Hieronym: Leveraging Hierarchical Multi-Source Information for Function Renaming in Stripped Binary
Function renaming in stripped binaries can substantially assist reverse engineers by improving code readability, yet it is a challenging task. The difficulty stems from the need to accurately capture function semantics from low-level binary code across diverse instruction sets, architectures, and...