Phying News
Curated security research, vulnerabilities, advisories and tools for practitioners.

Training State of the Art Vulnerability Discovery Agents through Reinforcement Learning

Summary

depthfirst introduces dfs-mini1, a security model co-trained with its harness via reinforcement learning; it reaches Pareto optimality on OpenAI's EVMBench Detect and state-of-the-art pass@8.
Published
Collected

original ↗

Related coverage

back