← Back to the trend map

Safety · Concept

Science of AI Evaluation

Joint Singapore-Korea AI Safety Institute evaluation of agent data leakage across 12 realistic non-adversarial tasks demonstrates no agent achieved fully safe execution.

Trend strength 7/10
Momentum +7/q
Confidence medium
Status new
Forecast horizon

Connections

Connections · 7

How this node ties into the rest of the map, and the evidence behind each link.

Signal sources

Signal sources

Dated facts from primary sources in this direction.