面向真实世界的渗透测试智能体评估
ethiack.com | 研究 | #ai-security | #pentesting | #agentic-ai | #benchmark | #llm | #ai-pentesting | #evaluation | #ethiack | #ethibench
摘要
现有的渗透测试基准在真实世界中并不奏效。了解 Ethiack 如何构建 EthiBench:一个在逼真目标上测试 AI 渗透测试智能体的评估框架,具备结构化真值、LLM 裁判匹配,以及超越夺旗的评估指标。
- 发布时间
- 收录时间
Skip to content