AI Briefing
KO

The Process of Building and Benchmarking the Frontier Cyber Reasoning Model VR-1

·2026.07.27 09:00

Key point

Cogent has released VR-1, a cyber reasoning model capable of autonomously executing attack chains, along with the IntrusionBench benchmark.

Details

Cogent's VR-1 is a frontier cyber reasoning model that can autonomously investigate environments, test hypotheses, and execute attack chains across system boundaries.

The newly released IntrusionBench is a benchmark that measures whether cyber agents can complete attack chains in real enterprise environments starting from only limited initial access.

In the black-box setting of IntrusionBench, VR-1 achieved more than a 2x improvement in pass@3 performance compared to the strongest existing frontier baseline model. However, since both VR-1 and IntrusionBench are currently in early preview stage, the results should be considered preliminary.

This summary was generated automatically by AI. Check the original for the author's claims and context. Copyright belongs to the original author.

Our guide explains how the AI works. Report summary errors, attribution issues, or removal requests via Contact.