← Back to the trend map

Safety · Concept

Virtue Ethics vs. Subordinate AI: Existential Risk Trade-off

Empirical study showing Constitutional AI fine-tuning for virtue reduces existential risk but increases general safety risks, revealing a fundamental alignment trade-off.

Trend strength 4/10
Momentum -1/q
Confidence medium
Status new
Forecast horizon

Trade-off between agent well-being and subordination will become central to long-term alignment research agendas.

Connections

Connections · 4

How this node ties into the rest of the map, and the evidence behind each link.

Signal sources

Signal sources

Dated facts from primary sources in this direction.