1. 介绍 dfbench v1 研究 | 2026-08-19 12:00 | depthfirst.com | 原文 ↗ | #benchmark | #ai-agents | #vulnerability-detection
2. Benchmaxxing:当基准测试本身成为攻击目标 博客 | 2026-08-19 00:00 | crowdstrike.com | 原文 ↗ | #ai-security | #benchmark | #cybersecurity
3. 面向真实世界的渗透测试智能体评估(第二部分) 研究 | 2026-07-22 00:00 | ethiack.com | 原文 ↗ | #ai-security | #pentesting | #agentic-ai
4. 面向真实世界的渗透测试智能体评估 研究 | 2026-07-14 00:00 | ethiack.com | 原文 ↗ | #ai-security | #pentesting | #agentic-ai
5. 跻身 Top 1%:打造 Tenzai 的 AI 黑客与精英人类竞争 研究 | 2026-06-25 00:00 | tenzai.com | 原文 ↗ | #ai-security | #benchmark | #ctf
6. 使用 agent-belt 让你的智能体尽在掌控 博客 | 2026-05-19 13:17 | jfrog.com | 原文 ↗ | #ai-security | #open-source | #mcp
7. Mythos 在攻防安全领域的表现:XBOW 的评测 研究 | 2026-05-12 12:00 | xbow.com | 原文 ↗ | #ai-security | #benchmark | #reverse-engineering
8. DEF CON 33:关于AI安全、AI红队和未来之路的现场笔记 博客 | 2025-08-21 14:03 | hackerone.com | 原文 ↗ | #ai-security | #def-con | #llm