We burned 11.7bn tokens to find the best cyber AI model
aikido.dev | blog | #ai-security | #open-source | #benchmark | #aikido | #llm-security | #vulnerability-discovery | #ai-models | #deepseek
Summary
Aikido burned 11.7B tokens testing 10 AI models against 32 fresh vulns: DeepSeek V4 Pro pooled the most finds (28/32), and open models now rival closed frontier ones—at the cost of noisier leads.
Why it matters
This coverage gives security teams current context for monitoring, validation, and remediation.
- Published
- Collected
Skip to content