← Back to the trend map

Safety · Concept

Threat-Model Coverage Gap in AI Safety Evaluation

Methodological critique that automated red-teaming gains on fixed benchmarks do not replace human evaluators with different deployment-context coverage.

Trend strength 5/10
Momentum +5/q
Confidence medium
Status new
Forecast horizon

Pushes evaluation science toward external, context-divergent evaluators rather than benchmark-only automation.

Connections

Connections · 4

How this node ties into the rest of the map, and the evidence behind each link.

Signal sources

Signal sources

Dated facts from primary sources in this direction.