Claude Fable 5 vs Qwen 3.6 Max (Preview)
Verdict
Claude Fable 5 leads 4–0 across 4 shared benchmarks.
Claude Fable 5 4 · Qwen 3.6 Max (Preview) 0 · higher isn't always better — see caveats.
Per-benchmark head-to-head
Reported scores — protocols may differ. How we compare
Sort
Claude Fable 5Qwen 3.6 Max (Preview)
Capabilities
Reasoning
Reasoning: Claude Fable 5 leads 2–0 across 2 shared reasoning benchmarks (largest gap: SimpleBench, 78.28 vs 55.6 — single evaluator).
2 reproduced2 unverified
What the labs claim
Vendor claims — compare against the measured scores above. A claim is what a developer says about its own model, not an independent measurement.
Claude Fable 5
Agentic / Tool Use
“The longer and more complex the task, the larger Fable 5's lead over our other models”
anthropic.com2026
Coding
“It is state-of-the-art on nearly all tested benchmarks of AI capability, showing exceptional performance in software engineering, knowledge work, vision, scientific research, and many other areas”
anthropic.com2026
General
“Fable 5's capabilities exceed those of any model we've ever made generally available”
anthropic.com2026
Multimodal
“It can extract precise numbers from detailed scientific figures and can perform complex vision-based tasks like rebuilding a web app's source code from screenshots alone”
anthropic.com2026
Qwen 3.6 Max (Preview)
Agentic / Tool Use
Coding
“It achieves the top score on six major coding benchmarks - SWE-bench Pro, Terminal-Bench 2.0, SkillsBench, QwenClawBench, QwenWebBench, and SciCode - with substantial gains over its predecessor”
qwen.ai2026