Safety · Trend
Narrowing Cyber Capability Gap of Open-Weight Models
AISI finds leading open-weight models trail frontier closed cyber capability by only 4–7 months, down from 6–10.
Open-weight cyber parity may compress defender lead times for evaluation and export controls.
Connections
Connections · 3
How this node ties into the rest of the map, and the evidence behind each link.
AISI cyber evaluation finds open-weight models only 4–7 months behind frontier closed models.
+4 growthLow RCA accuracy of frontier coding agents under realistic oncall noise complements cyber-capability gap measurements for operational risk.
+4 growthAISI cites DeepSeek V4-Pro among open models performing near frontier closed cyber capability.
+3 growthSignal sources
Signal sources
Dated facts from primary sources in this direction.
In June 2025 the US AI Safety Institute was renamed the Center for AI Standards and Innovation (CAISI), pivoting toward security, standards and adversary-model assessment.
NIST →Anthropic activated its ASL-3 deployment and security standard with Claude Opus 4 on 22 May 2025 — the first real-world trigger of a responsible-scaling tier, focused on blocking bio-weapon uplift.
Anthropic →The International Network of AI Safety Institutes (launched Nov 2024) ran a third joint testing exercise focused on agentic AI systems across cyber and fraud strands.
European Commission — AI Office →