llmboard.aiAI model intelligence
Home

Model Rankings

OverallOpen ModelsAgentCodingReasoningMathKnowledgeInstruction FollowingTextVision
Image GenerationImage Editing
Video GenerationImage to VideoVideo Editing
Text to SpeechSpeech to Text
Embeddings

Efficiency

Chat Token PricingImage PricingVideo PricingAudio Pricing
Chat Speed & LatencyProvider Reliability

Benchmarks

GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-Pro
All Benchmarks

Tools

Model DirectoryCompare Models

Scoring & Data

Scoring & Data
393 models668 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll BenchmarksReasoningMath

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai

OpenAI model product

o1

A research preview model focused on mathematical and logical reasoning capabilities, demonstrating improved performance on tasks requiring step-by-step reasoning, mathematical problem-solving, and code generation.

Updated Aug 17, 2026. Default version: o1

Compare
LLMBoard Score35.9o1
Coverage80%16 benchmark families
Context window200KTokens
Official input price$15OpenAI API

On this page

  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Runtime
  • Specification
  • Versions
  • Compare
  • Similar models
  • About
  • FAQ

o1 Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

o1 LLMBoard score breakdown

o1 Benchmark Results

Benchmark scores for o1.

19 rows
Columns

Show columns

Sort by
Benchmark
Score
Rank
Participants
Percentile
Evidence
Evaluated
BenchmarkGPQA BiologyScore69.2%Rank01Participants1Percentile100.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkGPQA ChemistryScore64.7%Rank01Participants1Percentile100.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkGPQA PhysicsScore92.8%Rank01Participants1Percentile100.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkMATHScore96.4%Rank02Participants71Percentile98.6%EvidenceCEvaluatedAug 17, 2026
BenchmarkMMLUScore91.8%Rank02Participants101Percentile99.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkGSM8kScore97.1%Rank03Participants48Percentile95.7%EvidenceCEvaluatedAug 17, 2026
BenchmarkMGSMScore89.3%Rank10Participants31Percentile70.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkTAU-bench RetailScore70.8%Rank10Participants25Percentile62.5%EvidenceCEvaluatedAug 17, 2026
BenchmarkMathVistaScore71.8%Rank12Participants39Percentile71.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkTAU-bench AirlineScore50.0%Rank12Participants23Percentile50.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkFrontierMathScore5.5%Rank17Participants17Percentile0.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkSimpleQAScore47.0%Rank17Participants47Percentile65.2%EvidenceCEvaluatedAug 17, 2026
BenchmarkMMMUScore77.6%Rank18Participants63Percentile72.6%EvidenceCEvaluatedAug 17, 2026
BenchmarkMMMLUScore87.7%Rank19Participants49Percentile62.5%EvidenceCEvaluatedAug 17, 2026
BenchmarkHumanEvalScore88.1%Rank26Participants66Percentile61.5%EvidenceCEvaluatedAug 17, 2026
BenchmarkLiveBenchScore67.0%Rank33Participants38Percentile13.5%EvidenceCEvaluatedAug 17, 2026
BenchmarkAIME 2024Score74.3%Rank35Participants53Percentile34.6%EvidenceCEvaluatedAug 17, 2026
BenchmarkGPQAScore78.0%Rank100Participants239Percentile58.4%EvidenceCEvaluatedAug 17, 2026
BenchmarkSWE-Bench VerifiedScore41.0%Rank101Participants111Percentile9.1%EvidenceCEvaluatedAug 17, 2026

o1 Arena Results

Preference and agent-evaluation results for the default version.

4 rows
Columns

Show columns

Sort by
Arena
Category
Rank
Rating / score
Votes
Observations
Result date
Arenavision style controlCategoryoverallRank76Rating / score1193.4Votes3,694ObservationsN/AResult dateAug 6, 2026
ArenavisionCategoryoverallRank89Rating / score1168.6Votes3,694ObservationsN/AResult dateAug 6, 2026
Arenatext style controlCategoryoverallRank136Rating / score1402.2Votes27,807ObservationsN/AResult dateAug 12, 2026
ArenatextCategoryoverallRank166Rating / score1365.9Votes27,807ObservationsN/AResult dateAug 12, 2026

o1 Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
$15 input, $60 output per 1M
Official provider
OpenAI
Lowest third-party
N/A
Tracked offerings
1
1 rows
Columns

Show columns

Sort by
Provider
Provider model ID
Region
Input / 1M
Output / 1M
Context
Updated
ProviderOpenAIProvider model IDo1RegionglobalInput / 1M$15Output / 1M$60Context200KUpdatedAug 17, 2026

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

o1 Runtime Performance

Provider-specific output speed and catalog latency for o1. Runtime does not affect the capability score.

2 rows
Columns

Show columns

Sort by
Provider
Output Speed
Catalog Latency
Max Input
Max Output
Updated
ProviderOpenAIOutput Speed66 tok/sCatalog Latency16.2 sMax Input200KMax Output100KUpdatedAug 17, 2026
ProviderAzureOutput Speed16 tok/sCatalog Latency0.54 sMax Input200KMax Output100KUpdatedAug 17, 2026

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

o1 Specifications

Technical details for the model's default version.

Version
o1
Released
Dec 17, 2024
Knowledge cutoff
Sep 1, 2023
Parameters
N/A
Context window
200K
Max output
100K
Inputs
image, pdf, text
Outputs
text
Open weights
No
License
Proprietary

o1 Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

2 rows
Columns

Show columns

Sort by
Version
Released
LLMBoard
Parameters
Context
Max output
Open weights
License
Versiono1ReleasedDec 17, 2024LLMBoard35.9ParametersN/AContext200KMax output100KOpen weightsNoLicenseProprietary
Versiono1-previewReleasedSep 12, 2024LLMBoardN/AParametersN/AContext128KMax output32.8KOpen weightsNoLicenseProprietary

o1 vs nearby models

Open a comparison with the three ranked models immediately above and below this model.

o1vsNorth Mini Code 1.0o1vsMiniMax M1 80Ko1vsLongCat Flash Chato1vsQwen3 Next 80B A3Bo1vsGPT-4.5o1vsGrok Code Fast 1

Models similar to o1

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#152-0.3
OP

GPT-4.5

OpenAI

35.7 LLMBoard

DetailsCompare
#154-0.7
OP

o3 mini

OpenAI

35.2 LLMBoard

DetailsCompare
#159-2.3
OP

o1 pro

OpenAI

33.6 LLMBoard

DetailsCompare
#137+3.8
OP

GPT-5.3

OpenAI

39.7 LLMBoard

DetailsCompare
#168-4.7
OP

GPT-4.1

OpenAI

31.3 LLMBoard

DetailsCompare
#171-5.6
OP

GPT-5-nano

OpenAI

30.3 LLMBoard

DetailsCompare

What is o1?

Key information about o1 and its available data.

A research preview model focused on mathematical and logical reasoning capabilities, demonstrating improved performance on tasks requiring step-by-step reasoning, mathematical problem-solving, and code generation. The model shows enhanced capabilities in formal reasoning while maintaining strong general capabilities.

Data as of 2026-08-17.

FAQ

Common questions about o1.

When was o1 released?

o1's default version was released on Dec 17, 2024.

How much does o1 cost?

o1's official API price is $15 per million input tokens and $60 per million output tokens via OpenAI.

Who created o1?

o1 was created by OpenAI.

What is the context window for o1?

The default version has a 200K token context window.

Is o1 open weight?

No. The default version is not marked as having publicly available weights.

How many API providers offer o1?

1 provider offerings are linked to the default version.

What models should I compare o1 with?

Nearby ranked alternatives include North Mini Code 1.0, MiniMax M1 80K, LongCat Flash Chat.