Evaluating Pentesting Agents for the Real-World
ethiack.com | research | #ai-security | #pentesting | #agentic-ai | #benchmark | #llm | #ai-pentesting | #evaluation | #ethiack | #ethibench
Summary
Ethiack launches EthiBench, an open evaluation protocol for AI pentesting agents using realistic targets, structured ground truth with LLM-as-a-judge matching, and richer metrics than success rates.
- Published
- Collected
Skip to content