1 matches found
Compared to What? A Human-Anchored Security Benchmark for LLM-Generated Infrastructure-As-Code
Large language models are increasingly used to author Infrastructure-as-Code IaC, where a single insecure default can be deployed directly into production. Prior evaluations report raw vulnerability counts for model-generated IaC, but without a human baseline they cannot determine whether models...
5.9AI score
SaveExploits0
20