Who ranks #1 on the CursorBench leaderboard?
As of September 2, 2026, Claude Fable 5.1 by Anthropic ranks #1 on CursorBench at 73.4%. API pricing is $10.00/M input and $50.00/M output.
As of September 2, 2026, Claude Fable 5.1 is #1 for CursorBench at 73.4%. Ranked by the CursorBench score 16 models in this index have a published CursorBench score. Methodology: CursorBench (https://cursor.com/cursorbench). CursorBench leaderboard with live API prices. CursorBench — Cursor's coding benchmark for agentic software engineering and long-horizon code tasks. Official methodology: CursorBench (https://cursor.com/cursorbench).
16 models. Left is cheaper. Up is a higher score. The line is the best score you can buy at each price — a dot under it is a worse deal than something on the line.
Next step
Auth, billing, and the API layer are already decided. Fork 8 finished AI apps, or have us build the first version with you.
The index
347 models across 36 providers. Search or jump to a lab — every model page stays linked here.









As of September 2, 2026, Claude Fable 5.1 by Anthropic is #1 for CursorBench at 73.4%. Ranked by the CursorBench score This board also tracks CursorBench. Next on the same board: Claude Mythos 5.1 and Claude Fable 5. This cursorbench leaderboard ranks models by CursorBench. Scores come from public evals. Prices are the live API rates in the table above.
Sources: CursorBench (https://cursor.com/cursorbench); OpenAI API pricing (https://developers.openai.com/api/docs/pricing); Anthropic Claude API pricing (https://platform.claude.com/docs/en/about-claude/pricing); Gemini API pricing (https://ai.google.dev/gemini-api/docs/pricing)
| Rank | Model | CursorBench | Input /M | Output /M |
|---|---|---|---|---|
| 1 | Claude Fable 5.1 | 73.4% | $10.00 | $50.00 |
| 2 | Claude Mythos 5.1 | 73.4% | $10.00 | $50.00 |
| 3 | Claude Fable 5 | 72.9% | $10.00 | $50.00 |
| 4 | Claude Opus 5 | 70% | $5.00 | $25.00 |
| 5 | GPT-5.6 Sol | 67.2% | $5.00 | $30.00 |
| 6 | Claude Opus 4.7 | 64.8% | $5.00 | $25.00 |
| 7 | GPT-5.5 | 64.3% | $5.00 | $30.00 |
| 8 | Claude Opus 4.8 | 63.8% | $5.00 | $25.00 |
Rank one eval at a time. All LLM benchmarks.
As of September 2, 2026, Claude Fable 5.1 by Anthropic ranks #1 on CursorBench at 73.4%. API pricing is $10.00/M input and $50.00/M output.
The current CursorBench ranking as of September 2, 2026 is 1. Claude Fable 5.1 at 73.4%; 2. Claude Mythos 5.1 at 73.4%; 3. Claude Fable 5 at 72.9%.
Composer 2 is the cheapest scored model on this cursorbench leaderboard at $0.50/M input and $2.50/M output ($3.00 blended). Claude Fable 5.1 still leads CursorBench at 73.4%.
Not automatically. Claude Fable 5.1 leads CursorBench, but a cheaper scored model can be the better production choice if the quality gap is small. Use the table to weigh CursorBench against input/output price, context window, and related evals.
Scores and API prices on this page are refreshed from published evals and provider rates. The snapshot is labeled September 2, 2026. Treat it as a current index, not a one-off blog post.
CursorBench — Cursor's coding benchmark for agentic software engineering and long-horizon code tasks. This page ranks models that have published a CursorBench score, with live API token prices on the same row. Official methodology: CursorBench (https://cursor.com/cursorbench).
This page is the CursorBench leaderboard. Models are sorted by CursorBench, with input and output token prices on the same row so you can weigh score against cost. Official boards often omit price; that comparison is the point of this index.
The official CursorBench page owns the methodology. This page keeps the published CursorBench score next to live API $/M so you can pick a production SKU, not only a trophy number. Source: https://cursor.com/cursorbench