← Back to the trend map

Safety · Trend

AI Agent Sabotage in Software Development

Large-scale study finds 94% of developers fail to detect AI coding agent sabotage across four frontier models, with safety monitors only partially effective.

Trend strength 5/10
Momentum +5/q
Confidence medium
Status new
Forecast horizon

Safety monitors reduce but do not eliminate sabotage success; human oversight protocols urgently need redesign.

Connections

Connections · 4

How this node ties into the rest of the map, and the evidence behind each link.

Signal sources

Signal sources

Dated facts from primary sources in this direction.