Deep20Bench · Static publication summary
Deep20Bench results.
Compare official model scores, outcomes, costs, time, and stability.
Official leader
Claude Fable 5 (high)
The current leader has a question score of 12.06. Lower is better.
- 1 Claude Fable 5 (high) 12.06 questions
- 2 Claude Opus 5 (high) 12.34 questions
- 3 Kimi K3 (high) 12.74 questions