Capability ranking
Use this reasoning model leaderboard to compare the current ranking from the LLMBoard Reasoning Score after applying benchmark evidence requirements.
Updated 2026-08-17
This leaderboard ranking orders models by their LLMBoard Reasoning Score aggregated from eligible benchmark evidence.
Rank | Model | LLMBoard Reasoning Score | Context | Official input / 1M | Official output / 1M |
|---|
| Rank01 | ModelAN | LLMBoard Reasoning Score99.1 | Context1M | Official input / 1M$5 | Official output / 1M$25 |
| Rank02 | ModelOP | LLMBoard Reasoning Score99.1 | Context1.1M | Official input / 1M$5 | Official output / 1M$30 |
| Rank03 | ModelAC | LLMBoard Reasoning Score97.7 | Context1M | Official input / 1M$2 | Official output / 1M$6 |
| Rank04 | ModelGO | LLMBoard Reasoning Score96.5 | Context1M | Official input / 1M$2 | Official output / 1M$12 |
| Rank05 | ModelOP | LLMBoard Reasoning Score91.7 | Context1.1M | Official input / 1M$2.5 | Official output / 1M$15 |
| Rank06 | ModelAC | LLMBoard Reasoning Score91.6 | Context1M | Official input / 1M$2.5 | Official output / 1M$7.5 |
| Rank07 | ModelBY | LLMBoard Reasoning Score90.1 | Context256K | Official input / 1MN/A | Official output / 1MN/A |
| Rank08 | ModelAC | LLMBoard Reasoning Score87.9 | Context1M | Official input / 1M$0.50 | Official output / 1M$3 |
| Rank09 | ModelGO | LLMBoard Reasoning Score87.0 | Context262.1K | Official input / 1MN/A | Official output / 1MN/A |
| Rank10 | ModelAN | LLMBoard Reasoning Score84.1 | Context1M | Official input / 1M$5 | Official output / 1M$25 |
| Rank11 | ModelXI | LLMBoard Reasoning Score83.9 | Context1M | Official input / 1M$0.435 | Official output / 1M$0.87 |
| Rank12 | ModelAC | LLMBoard Reasoning Score81.3 | Context262.1K | Official input / 1MN/A | Official output / 1MN/A |
| Rank13 | ModelAC | LLMBoard Reasoning Score81.0 | Context262.1K | Official input / 1M$0.60 | Official output / 1M$3.6 |
| Rank14 | ModelOP | LLMBoard Reasoning Score79.8 | Context400K | Official input / 1M$21 | Official output / 1M$168 |
| Rank15 | ModelAN | LLMBoard Reasoning Score79.3 | Context1M | Official input / 1M$3 | Official output / 1M$15 |
| Rank16 | ModelGO | LLMBoard Reasoning Score78.7 | Context262.1K | Official input / 1MN/A | Official output / 1MN/A |
| Rank17 | ModelAN | LLMBoard Reasoning Score77.6 | Context200K | Official input / 1MN/A | Official output / 1MN/A |
| Rank18 | ModelBY | LLMBoard Reasoning Score76.4 | Context256K | Official input / 1MN/A | Official output / 1MN/A |
| Rank19 | ModelMA | LLMBoard Reasoning Score76.2 | Context262.1K | Official input / 1M$0.60 | Official output / 1M$3 |
| Rank20 | ModelAC | LLMBoard Reasoning Score75.7 | Context1M | Official input / 1M$0.50 | Official output / 1M$3 |
| Rank21 | ModelOP | LLMBoard Reasoning Score75.5 | Context400K | Official input / 1M$1.75 | Official output / 1M$14 |
| Rank22 | ModelAN | LLMBoard Reasoning Score73.7 | Context200K | Official input / 1MN/A | Official output / 1MN/A |
| Rank23 | ModelGO | LLMBoard Reasoning Score72.9 | Context1M | Official input / 1MN/A | Official output / 1MN/A |
| Rank24 | ModelGO | LLMBoard Reasoning Score71.5 | Context1M | Official input / 1M$0.50 | Official output / 1M$3 |
| Rank25 | ModelCO | LLMBoard Reasoning Score71.2 | Context128K | Official input / 1M$2.5 | Official output / 1M$10 |
| Rank26 | ModelAC | LLMBoard Reasoning Score70.9 | Context262.1K | Official input / 1MN/A | Official output / 1MN/A |
| Rank27 | ModelME | LLMBoard Reasoning Score70.4 | ContextN/A | Official input / 1MN/A | Official output / 1MN/A |
| Rank28 | ModelGO | LLMBoard Reasoning Score69.8 | ContextN/A | Official input / 1MN/A | Official output / 1MN/A |
| Rank29 | ModelGO | LLMBoard Reasoning Score68.6 | ContextN/A | Official input / 1MN/A | Official output / 1MN/A |
| Rank30 | ModelAM | LLMBoard Reasoning Score68.5 | Context300K | Official input / 1M$0.80 | Official output / 1M$3.2 |
| Rank31 | ModelOP | LLMBoard Reasoning Score67.5 | Context8.2K | Official input / 1M$30 | Official output / 1M$60 |
| Rank32 | ModelME | LLMBoard Reasoning Score67.0 | Context128K | Official input / 1MN/A | Official output / 1MN/A |
| Rank33 | ModelTM | LLMBoard Reasoning Score66.5 | Context1M | Official input / 1MN/A | Official output / 1MN/A |
| Rank34 | ModelAC | LLMBoard Reasoning Score62.4 | Context262.1K | Official input / 1M$0.40 | Official output / 1M$3.2 |
| Rank35 | ModelGO | LLMBoard Reasoning Score61.9 | Context2.1M | Official input / 1MN/A | Official output / 1MN/A |
| Rank36 | ModelMI | LLMBoard Reasoning Score61.7 | ContextN/A | Official input / 1MN/A | Official output / 1MN/A |
| Rank37 | ModelAC | LLMBoard Reasoning Score60.7 | Context262.1K | Official input / 1M$0.70 | Official output / 1M$2.8 |
| Rank38 | ModelAN | LLMBoard Reasoning Score59.4 | Context200K | Official input / 1M$5 | Official output / 1M$25 |
| Rank39 | ModelGO | LLMBoard Reasoning Score59.4 | ContextN/A | Official input / 1MN/A | Official output / 1MN/A |
| Rank40 | ModelNR | LLMBoard Reasoning Score59.1 | Context131.1K | Official input / 1MN/A | Official output / 1MN/A |
| Rank41 | ModelAN | LLMBoard Reasoning Score58.3 | Context200K | Official input / 1MN/A | Official output / 1MN/A |
| Rank42 | ModelAC | LLMBoard Reasoning Score57.8 | Context131.1K | Official input / 1M$0.50 | Official output / 1M$6 |
| Rank43 | ModelNV | LLMBoard Reasoning Score57.3 | Context128K | Official input / 1MN/A | Official output / 1MN/A |
| Rank44 | ModelAC | LLMBoard Reasoning Score54.9 | Context262.1K | Official input / 1M$0.248 | Official output / 1M$1.49 |
| Rank45 | ModelGO | LLMBoard Reasoning Score54.9 | Context1M | Official input / 1MN/A | Official output / 1MN/A |
| Rank46 | ModelOP | LLMBoard Reasoning Score53.6 | Context128K | Official input / 1M$10 | Official output / 1M$30 |
| Rank47 | ModelAC | LLMBoard Reasoning Score53.2 | Context262.1K | Official input / 1MN/A | Official output / 1MN/A |
| Rank48 | ModelAC | LLMBoard Reasoning Score53.2 | Context262.1K | Official input / 1M$0.25 | Official output / 1M$2 |
| Rank49 | ModelAC | LLMBoard Reasoning Score52.9 | Context131.1K | Official input / 1M$0.50 | Official output / 1M$2 |
| Rank50 | ModelME | LLMBoard Reasoning Score52.8 | Context128K | Official input / 1MN/A | Official output / 1MN/A |
| Rank51 | ModelAC | LLMBoard Reasoning Score51.8 | Context262.1K | Official input / 1M$0.60 | Official output / 1M$3.6 |
| Rank52 | ModelAM | LLMBoard Reasoning Score51.3 | Context300K | Official input / 1M$0.06 | Official output / 1M$0.24 |
| Rank53 | ModelAC | LLMBoard Reasoning Score51.2 | Context262.1K | Official input / 1MN/A | Official output / 1MN/A |
| Rank54 | ModelAC | LLMBoard Reasoning Score50.8 | Context262.1K | Official input / 1M$0.30 | Official output / 1M$2.4 |
| Rank55 | ModelAC | LLMBoard Reasoning Score50.8 | ContextN/A | Official input / 1MN/A | Official output / 1MN/A |
| Rank56 | ModelXA | LLMBoard Reasoning Score50.8 | Context256K | Official input / 1MN/A | Official output / 1MN/A |
| Rank57 | ModelAC | LLMBoard Reasoning Score50.5 | ContextN/A | Official input / 1M$0.70 | Official output / 1M$2.8 |
| Rank58 | ModelGO | LLMBoard Reasoning Score49.9 | ContextN/A | Official input / 1MN/A | Official output / 1MN/A |
| Rank59 | ModelME | LLMBoard Reasoning Score48.7 | Context128K | Official input / 1MN/A | Official output / 1MN/A |
| Rank60 | ModelAC | LLMBoard Reasoning Score47.3 | Context262.1K | Official input / 1MN/A | Official output / 1MN/A |
| Rank61 | ModelAC | LLMBoard Reasoning Score46.9 | ContextN/A | Official input / 1MN/A | Official output / 1MN/A |
| Rank62 | ModelMA | LLMBoard Reasoning Score46.9 | Context128K | Official input / 1M$0.15 | Official output / 1M$0.15 |
| Rank63 | ModelGO | LLMBoard Reasoning Score46.6 | Context131.1K | Official input / 1MN/A | Official output / 1MN/A |
| Rank64 | ModelAN | LLMBoard Reasoning Score46.4 | Context200K | Official input / 1MN/A | Official output / 1MN/A |
| Rank65 | ModelAC | LLMBoard Reasoning Score46.3 | Context131.1K | Official input / 1MN/A | Official output / 1MN/A |
| Rank66 | ModelAL | LLMBoard Reasoning Score46.0 | Context256K | Official input / 1MN/A | Official output / 1MN/A |
| Rank67 | ModelOP | LLMBoard Reasoning Score45.9 | Context200K | Official input / 1M$2 | Official output / 1M$8 |
| Rank68 | ModelGO | LLMBoard Reasoning Score44.9 | Context131.1K | Official input / 1MN/A | Official output / 1MN/A |
| Rank69 | ModelAN | LLMBoard Reasoning Score42.1 | Context200K | Official input / 1MN/A | Official output / 1MN/A |
| Rank70 | ModelAM | LLMBoard Reasoning Score40.0 | Context128K | Official input / 1M$0.035 | Official output / 1M$0.14 |
| Rank71 | ModelGO | LLMBoard Reasoning Score39.3 | Context131.1K | Official input / 1MN/A | Official output / 1MN/A |
| Rank72 | ModelAC | LLMBoard Reasoning Score38.5 | ContextN/A | Official input / 1MN/A | Official output / 1MN/A |
| Rank73 | ModelAC | LLMBoard Reasoning Score37.3 | Context262.1K | Official input / 1MN/A | Official output / 1MN/A |
| Rank74 | ModelAN | LLMBoard Reasoning Score36.8 | Context200K | Official input / 1MN/A | Official output / 1MN/A |
| Rank75 | ModelAC | LLMBoard Reasoning Score35.6 | ContextN/A | Official input / 1MN/A | Official output / 1MN/A |
| Rank76 | ModelOP | LLMBoard Reasoning Score35.2 | Context128K | Official input / 1M$0.15 | Official output / 1M$0.60 |
| Rank77 | ModelAL | LLMBoard Reasoning Score34.8 | Context256.1K | Official input / 1MN/A | Official output / 1MN/A |
| Rank78 | ModelMA | LLMBoard Reasoning Score34.1 | Context131.1K | Official input / 1MN/A | Official output / 1MN/A |
| Rank79 | ModelMI | LLMBoard Reasoning Score33.5 | Context16K | Official input / 1M$0.125 | Official output / 1M$0.50 |
| Rank80 | ModelGO | LLMBoard Reasoning Score32.8 | Context131.1K | Official input / 1MN/A | Official output / 1MN/A |
| Rank81 | ModelAC | LLMBoard Reasoning Score31.8 | Context262.1K | Official input / 1MN/A | Official output / 1MN/A |
| Rank82 | ModelAC | LLMBoard Reasoning Score30.1 | Context262.1K | Official input / 1MN/A | Official output / 1MN/A |
| Rank83 | ModelAC | LLMBoard Reasoning Score27.8 | Context262.1K | Official input / 1MN/A | Official output / 1MN/A |
| Rank84 | ModelMI | LLMBoard Reasoning Score27.6 | Context128K | Official input / 1MN/A | Official output / 1MN/A |
| Rank85 | ModelIB | LLMBoard Reasoning Score27.4 | Context128K | Official input / 1MN/A | Official output / 1MN/A |
| Rank86 | ModelME | LLMBoard Reasoning Score26.4 | Context128K | Official input / 1MN/A | Official output / 1MN/A |
| Rank87 | ModelGO | LLMBoard Reasoning Score25.7 | ContextN/A | Official input / 1MN/A | Official output / 1MN/A |
| Rank88 | ModelMI | LLMBoard Reasoning Score23.8 | Context128K | Official input / 1M$0.075 | Official output / 1M$0.30 |
| Rank89 | ModelME | LLMBoard Reasoning Score23.2 | Context128K | Official input / 1MN/A | Official output / 1MN/A |
| Rank90 | ModelAC | LLMBoard Reasoning Score22.8 | ContextN/A | Official input / 1M$0.144 | Official output / 1M$0.287 |
| Rank91 | ModelOP | LLMBoard Reasoning Score22.3 | Context128K | Official input / 1M$2.5 | Official output / 1M$10 |
| Rank92 | ModelAC | LLMBoard Reasoning Score22.0 | Context262.1K | Official input / 1MN/A | Official output / 1MN/A |
| Rank93 | ModelGO | LLMBoard Reasoning Score21.7 | Context131.1K | Official input / 1MN/A | Official output / 1MN/A |
| Rank94 | ModelAC | LLMBoard Reasoning Score20.2 | ContextN/A | Official input / 1M$0.35 | Official output / 1M$1.4 |
| Rank95 | ModelOP | LLMBoard Reasoning Score18.0 | Context16.4K | Official input / 1M$0.50 | Official output / 1M$1.5 |
| Rank96 | ModelGO | LLMBoard Reasoning Score16.2 | ContextN/A | Official input / 1MN/A | Official output / 1MN/A |
| Rank97 | ModelIB | LLMBoard Reasoning Score15.9 | ContextN/A | Official input / 1MN/A | Official output / 1MN/A |
| Rank98 | ModelAC | LLMBoard Reasoning Score14.9 | ContextN/A | Official input / 1MN/A | Official output / 1MN/A |
| Rank99 | ModelAC | LLMBoard Reasoning Score12.2 | Context262.1K | Official input / 1MN/A | Official output / 1MN/A |
| Rank100 | ModelGO | LLMBoard Reasoning Score9.5 | ContextN/A | Official input / 1MN/A | Official output / 1MN/A |
| Rank101 | ModelGO | LLMBoard Reasoning Score8.0 | Context32K | Official input / 1MN/A | Official output / 1MN/A |
| Rank102 | ModelBA | LLMBoard Reasoning Score6.3 | Context128K | Official input / 1MN/A | Official output / 1MN/A |
| Rank103 | ModelAC | LLMBoard Reasoning Score2.2 | ContextN/A | Official input / 1MN/A | Official output / 1MN/A |
| Rank104 | ModelGO | LLMBoard Reasoning Score1.8 | ContextN/A | Official input / 1MN/A | Official output / 1MN/A |
| Rank105 | ModelGO | LLMBoard Reasoning Score0.1 | ContextN/A | Official input / 1MN/A | Official output / 1MN/A |
Official PAYG prices appear here. Third-party offers remain on the pricing and model detail pages.
Key findings from the current reasoning leaderboard ranking and its supporting benchmark coverage.
Claude Opus 4.8 leads this page at 99.1, 0.0 points ahead of GPT-5.5.
Gemma 4 31B is the highest-ranked open-weight option at #9. Nova Micro has the lowest official input price among ranked models.
The leading reasoning scores in this benchmark-backed ranking, with the axis focused on the competitive leaderboard range.
Compare reasoning benchmark strength with each model's broader capability score and leaderboard position.
Official vendor API prices plotted against the reasoning metric used on this page. Price remains a separate decision signal.
Vendor concentration and model access among the first 25 products in the ranking.
Open focused comparisons between the current leader and the nearest practical alternatives.
A data-backed look at the first five models in this reasoning leaderboard ranking, including benchmark context, price and output speed where available.
Ranking basisThis reasoning AI model leaderboard uses the reasoning capability score shown on this page. The leaderboard ranking keeps matched price and speed data separate from benchmark evidence.
Selection summary
Claude Opus 4.8 is currently the best-ranked LLM for reasoning with a 99.1 reasoning score. The score is a relative ranking signal, so price, speed, context and evidence coverage should still be checked separately.
Use this leaderboard with the supporting benchmark results and coverage details above. A leaderboard position summarizes the selected ranking signal; it does not replace workload-specific testing.
Common questions about the Best AI for Reasoning leaderboard, benchmark evidence and ranking method.
Claude Opus 4.8 is currently ranked first with a reasoning score of 99.1.
Each model appears once using its current scored version. The leaderboard ranking follows the capability named in the title, while the overall page uses the LLMBoard score aggregated from eligible benchmark evidence.
This page currently ranks 105 unique model products.
Main price columns use the model vendor's official standard PAYG API rate. Eligible third-party offers appear only in separately labeled columns, and unavailable official prices display as N/A.
No. Arena results are displayed as an independent signal and are not included in the current LLMBoard capability score.