1 matches found
Big Enough to Break Out: Tracking the Rising Capability of LLM Penetration-Testing Agents
Large language model LLM agents are increasingly applied to penetration testing, but we still know little about what they can do or how they fail. We compare two PentestGPT-based systems: a legacy human-in-the-loop system running the open-weight Kimi K2.5, and a newer autonomous system running...
5.8AI score
SaveExploits0
20