← Back to the trend map

Capabilities · Trend

LLM Reasoning as Pattern Matching

DeFAb benchmark shows frontier LLMs reach only 65% accuracy on defeasible abduction tasks solvable by rule-based solvers in under 50 microseconds, confirming reasoning limitations.

Trend strength 6/10
Momentum +6/q
Confidence medium
Status new
Forecast horizon

Shared human-LLM reasoning failure modes challenge assumptions about abstract world models in both AI and cognitive science.

Connections

Connections · 2

How this node ties into the rest of the map, and the evidence behind each link.

Signal sources

Signal sources

Dated facts from primary sources in this direction.