Evaluating Pentesting Agents for the Real-World Part 2
ethiack.com | research | #ai-security | #pentesting | #agentic-ai | #benchmark | #llm | #security-testing | #ai-pentesting | #evaluation | #ethiack
Summary
Ethiack's EthiBench part 2 evaluates AI pentesting agents beyond single runs, covering significance under budget limits, cumulative detection, temporal behavior, and cost-effective target subsets.
- Published
- Collected
Skip to content