llmboard.aiLeaderboard Center
Overall
Overall RankingOpen Models
Tools
Model DirectoryCompare Models
Capabilities
CodingReasoningMathKnowledgeInstruction Following
Price & Efficiency
Price & ValueCapability vs. PriceRuntime Performance
Modalities
Image GenerationVideo GenerationSpeech ModelsEmbeddings
Core Benchmarks
GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-ProView all benchmarks
Methods
Scoring & Data
393 models668 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll BenchmarksReasoningMath

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai

Microsoft model product

Phi 4

phi-4 is a state-of-the-art open model built to excel at advanced reasoning, coding, and knowledge tasks.

Updated Aug 12, 2026. Default version: Phi 4

Compare
LLMBoard score9.4Phi 4
Coverage100%13 benchmark families
Context window16KTokens
Official input price$0.125Azure API

On this page

  • Specification
  • Capability
  • Benchmarks
  • Arena
  • Runtime
  • Pricing
  • Versions
  • About
  • Compare
  • Similar models
  • FAQ

Phi 4 Specifications

Technical details for the model's default version.

Version
Phi 4
Released
Dec 12, 2024
Knowledge cutoff
Jun 1, 2024
Parameters
14.7B
Context window
16K
Max output
16K
Inputs
text
Outputs
text
Open weights
No
License
MIT

Phi 4 Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

Phi 4 category scores

Phi 4 Benchmark Results

Benchmark scores for Phi 4.

13 rows
Columns

Show columns

PhiBench56.2%0330.0%CAug 11, 2026
HumanEval+82.8%051055.6%CAug 11, 2026
Arena Hard75.4%092668.0%CAug 11, 2026
DROP75.5%193037.9%CAug 11, 2026
MATH80.4%197174.3%CAug 11, 2026
MGSM80.6%193140.0%CAug 11, 2026
LiveBench47.6%36385.4%CAug 11, 2026
HumanEval82.6%416638.5%CAug 11, 2026
SimpleQA3.0%44464.4%CAug 11, 2026
MMLU84.8%4510055.6%CAug 11, 2026
IFEval63.0%63653.1%CAug 11, 2026
MMLU-Pro70.4%8212936.7%CAug 11, 2026
GPQA56.1%16523429.6%CAug 11, 2026

Phi 4 Arena Results

Preference and agent-evaluation results for the default version.

2 rows
Columns

Show columns

textoverall2841216.824,126N/AAug 10, 2026
text style controloverall2901256.224,126N/AAug 10, 2026

Phi 4 Runtime Performance

Provider-specific output speed and catalog latency for Phi 4. Runtime does not affect the capability score.

1 rows
Columns

Show columns

DeepInfra33 tok/s0.2 s16K16KAug 11, 2026

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

Phi 4 Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
$0.125 input, $0.50 output per 1M
Official provider
Azure
Lowest third-party
From $0.07 input, $0.14 output per 1M via OpenRouter
Tracked offerings
3
3 rows
Columns

Show columns

OpenRoutermicrosoft/phi-4global$0.07$0.1416.4KAug 11, 2026
Azure Cognitive Servicesphi-4global$0.125$0.50128KAug 11, 2026
Azurephi-4global$0.125$0.50128KAug 11, 2026

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

Phi 4 Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

1 rows
Columns

Show columns

Phi 4Dec 12, 20249.414.7B16K16KNoMIT

What is Phi 4?

Key information about Phi 4 and its available data.

phi-4 is a state-of-the-art open model built to excel at advanced reasoning, coding, and knowledge tasks. It leverages a blend of synthetic data, filtered web data, academic texts, and supervised fine-tuning for precision, alignment, and safety.

Data as of 2026-08-11.

Phi 4 vs nearby models

Open a comparison with the three ranked models immediately above and below this model.

Phi 4vsLlama 3.1 Nemotron Nano 8BPhi 4vsQwen2.5 Omni 7BPhi 4vsQwen2 72BPhi 4vsClaude Haiku 3.5Phi 4vsMistral Small 3.1 24BPhi 4vsQwen2.5 14B

Models similar to Phi 4

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#229-1.9
MI

Phi 4 multimodal

Microsoft

7.4 LLMBoard

DetailsCompare
#211+4.5
MI

Phi 4 Mini Reasoning

Microsoft

13.8 LLMBoard

DetailsCompare
#233-4.5
MI

Phi 3.5 MoE

Microsoft

4.9 LLMBoard

DetailsCompare
#251-9.4
MI

Phi 3.5 vision

Microsoft

0.0 LLMBoard

DetailsCompare
#256-9.4
MI

Phi 3.5 mini

Microsoft

0.0 LLMBoard

DetailsCompare
#277-9.4
MI

Phi 4 Mini

Microsoft

0.0 LLMBoard

DetailsCompare

FAQ

Common questions about Phi 4.

When was Phi 4 released?

Phi 4's default version was released on Dec 12, 2024.

How much does Phi 4 cost?

Phi 4's official API price is $0.125 per million input tokens and $0.50 per million output tokens via Azure. The lowest tracked third-party offer starts at $0.07 input and $0.14 output via OpenRouter.

Who created Phi 4?

Phi 4 was created by Microsoft.

What is the context window for Phi 4?

The default version has a 16K token context window.

Is Phi 4 open weight?

No. The default version is not marked as having publicly available weights.

How many API providers offer Phi 4?

3 provider offerings are linked to the default version.

What models should I compare Phi 4 with?

Nearby ranked alternatives include Llama 3.1 Nemotron Nano 8B, Qwen2.5 Omni 7B, Qwen2 72B.