Who ranks #1 on the AA-Briefcase leaderboard?
As of September 2, 2026, Claude Fable 5.1 by Anthropic ranks #1 on AA-Briefcase at 1694. API pricing is $10.00/M input and $50.00/M output.
As of September 2, 2026, Claude Fable 5.1 is #1 for AA-Briefcase at 1694. Ranked by the AA-Briefcase score 5 models in this index have a published AA-Briefcase score. Methodology: Artificial Analysis (https://artificialanalysis.ai/). AA-Briefcase leaderboard with live API prices. AA-Briefcase — Artificial Analysis Elo on multi-week knowledge-work projects. Official methodology: Artificial Analysis (https://artificialanalysis.ai/).
Next step
Auth, billing, and the API layer are already decided. Fork 8 finished AI apps, or have us build the first version with you.
The index
347 models across 36 providers. Search or jump to a lab — every model page stays linked here.









As of September 2, 2026, Claude Fable 5.1 by Anthropic is #1 for AA-Briefcase at 1694. Ranked by the AA-Briefcase score This board also tracks AA-Briefcase. Next on the same board: Claude Mythos 5.1 and Grok 4.6. This aa-briefcase leaderboard ranks models by AA-Briefcase. Scores come from public evals. Prices are the live API rates in the table above.
Sources: Artificial Analysis (https://artificialanalysis.ai/); OpenAI API pricing (https://developers.openai.com/api/docs/pricing); Anthropic Claude API pricing (https://platform.claude.com/docs/en/about-claude/pricing); Gemini API pricing (https://ai.google.dev/gemini-api/docs/pricing)
| Rank | Model | AA-Briefcase | Input /M | Output /M |
|---|---|---|---|---|
| 1 | Claude Fable 5.1 | 1694 | $10.00 | $50.00 |
| 2 | Claude Mythos 5.1 | 1694 | $10.00 | $50.00 |
| 3 | Grok 4.6 | 52.6% | $2.00 | $6.00 |
| 4 | Kimi K3 | 51.6% | $3.00 | $15.00 |
| 5 | Inkling-Small | 30.6% | $0.30 | $1.20 |
Rank one eval at a time. All LLM benchmarks.
As of September 2, 2026, Claude Fable 5.1 by Anthropic ranks #1 on AA-Briefcase at 1694. API pricing is $10.00/M input and $50.00/M output.
The current AA-Briefcase ranking as of September 2, 2026 is 1. Claude Fable 5.1 at 1694; 2. Claude Mythos 5.1 at 1694; 3. Grok 4.6 at 52.6%.
Inkling-Small is the cheapest scored model on this aa-briefcase leaderboard at $0.30/M input and $1.20/M output ($1.50 blended). Claude Fable 5.1 still leads AA-Briefcase at 1694.
Not automatically. Claude Fable 5.1 leads AA-Briefcase, but a cheaper scored model can be the better production choice if the quality gap is small. Use the table to weigh AA-Briefcase against input/output price, context window, and related evals.
Scores and API prices on this page are refreshed from published evals and provider rates. The snapshot is labeled September 2, 2026. Treat it as a current index, not a one-off blog post.
AA-Briefcase — Artificial Analysis Elo on multi-week knowledge-work projects. This page ranks models that have published a AA-Briefcase score, with live API token prices on the same row. Official methodology: Artificial Analysis (https://artificialanalysis.ai/).
This page is the AA-Briefcase leaderboard. Models are sorted by AA-Briefcase, with input and output token prices on the same row so you can weigh score against cost. Official boards often omit price; that comparison is the point of this index.
The official Artificial Analysis page owns the methodology. This page keeps the published AA-Briefcase score next to live API $/M so you can pick a production SKU, not only a trophy number. Source: https://artificialanalysis.ai/