Safety · Concept
Affective AI Safety
Safety concerns for emotionally responsive companion and affective models, including attachment and self-harm-related internal states.
Dedicated regulatory frameworks for affective harms may emerge as AI emotional engagement deepens.
Connections
Connections · 4
How this node ties into the rest of the map, and the evidence behind each link.
Valence-first ethics reframes affective/sentience safety away from unresolvable consciousness claims.
+4 growthAffective AI safety requires dedicated evaluation frameworks for cumulative, relational, and identity-level harms.
+3 growthEmotionally salient long conversations can induce ontological shock, expanding affective safety concerns.
+3 growthAffective safety harms are cumulative and relational, making them poorly captured by existing evaluation frameworks.
+2 growthSignal sources
Signal sources
Dated facts from primary sources in this direction.
In June 2025 the US AI Safety Institute was renamed the Center for AI Standards and Innovation (CAISI), pivoting toward security, standards and adversary-model assessment.
NIST →Anthropic activated its ASL-3 deployment and security standard with Claude Opus 4 on 22 May 2025 — the first real-world trigger of a responsible-scaling tier, focused on blocking bio-weapon uplift.
Anthropic →The International Network of AI Safety Institutes (launched Nov 2024) ran a third joint testing exercise focused on agentic AI systems across cyber and fraud strands.
European Commission — AI Office →