Post #8041 141 Sep 1, 2026, 08:00 UTC 🔗Ссылка:https://projectdiscovery.io/blog/watching-agents-work-a-behavioral-audit-of-offensive-security-llm-runs ProjectDiscovery Watching Agents Work: A Behavioral Audit of Offensive-Security LLM Runs — ProjectDiscovery Blog What closed and open models actually do when you tell them to hack a website Summary The cybersecurity capability of a model is currently measured by a solve rate, a percentage, and that number tells you almost nothing worth knowing. It doesn't tell you…