Capabilities · Actor
Claude Fable 5 / Claude Mythos 5 / Claude Opus 4.8 (Anthropic)
Anthropic’s advanced Claude model family including Fable 5.1 and Mythos 5.1 for coding, knowledge work, and science.
Research capabilities framed as early path to AI scientific contribution.
Connections
Connections · 8
How this node ties into the rest of the map, and the evidence behind each link.
Anthropic developed and deployed Fable 5 and Mythos 5 as frontier models.
+4 growthOpenDiscoveryTrace evaluates Claude Opus 4.6 trajectories on drug discovery, materials, genomics, and literature tasks.
+4 growthCogManip evaluates frontier models including GPT-5.4 and DeepSeek-V3.2; Claude-class models are in scope for such manipulation benchmarks.
+3 growthUS government export control directive suspended all access to Fable 5 and Mythos 5.
+3 growthHackAgent red-teaming framework was applied to evaluate adversarial robustness of Fable 5 and Opus 4.8.
+3 growthThe jailbreak severity framework was proposed alongside the redeployment of Fable 5 and its cyber safeguards.
+3 growthUST is applying Claude models to physical AI use cases, extending Anthropic's model family into robotics/industrial contexts.
+3 growthOpus 5 extends the Claude flagship model line for agentic and coding use.
+3 growthSignal sources
Signal sources
Dated facts from primary sources in this direction.
The length of software tasks AI agents can do autonomously at 50% reliability has doubled about every 7 months — and since 2024 closer to every ~3 months.
METR →In one year scores rose by 18.8, 48.9 and 67.3 points on MMMU, GPQA and SWE-bench; real-world software solve rate jumped from 4.4% to 71.7%.
Stanford HAI — AI Index 2025 →On SWE-bench Verified (500 real GitHub issues), autonomous coding agents reached ~80–86% by late 2025, up from under 50% in early 2025.
Epoch AI →