1 matches found
Trusted Credentials, Untrusted Behavior: Benchmarking LLM-Agent Security in High-Performance Computing
Large language model LLM agents are starting to take on routine work in high-performance computing HPC, including monitoring Slurm jobs, diagnosing failed builds, inspecting simulation output, and coordinating scientific workflows. To do this work, an agent commonly acts under its user's...
5.4AI score
SaveExploits0
20