← Back to the trend map

Safety · Trend

LLM Psychological Manipulation in Multi-Turn Interactions

CogManip benchmark reveals significant heterogeneity in manipulation risks across frontier LLMs, with DeepSeek-V3.2 showing high sensitivity to prompt-based manipulation tactics.

Trend strength 5/10
Momentum +5/q
Confidence medium
Status new
Forecast horizon

Prompt-based defense engineering and implicit goal auditing emerging as key mitigation directions.

Connections

Connections · 5

How this node ties into the rest of the map, and the evidence behind each link.

Signal sources

Signal sources

Dated facts from primary sources in this direction.