- Published on
A combined CursorBench and DeepSWE leaderboard comparing Opus 5, GPT-5.6, Kimi K3, Fable 5, and other coding models on correctness and cost.
AI coding benchmark analysis and leaderboards, comparing models on correctness and cost to help you pick a model stack.