llmboard.aiLeaderboard Center
Overall
Overall RankingOpen Models
Tools
Model DirectoryCompare Models
Capabilities
CodingReasoningMathKnowledgeInstruction Following
Price & Efficiency
Price & ValueCapability vs. PriceRuntime Performance
Modalities
Image GenerationVideo GenerationSpeech ModelsEmbeddings
Core Benchmarks
GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-ProView all benchmarks
Methods
Scoring & Data
393 models668 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll BenchmarksReasoningMath

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai

Microsoft model product

Phi 3.5 vision

2B-parameter open multimodal model with up to 128K context tokens.

Updated Aug 12, 2026. Default version: Phi-3.5-vision-instruct

Compare
LLMBoard score0.0Phi-3.5-vision-instruct
Coverage0%6 benchmark families
Context windowN/ATokens
Official input priceN/AOfficial price unavailable

On this page

  • Specification
  • Capability
  • Benchmarks
  • Arena
  • Runtime
  • Pricing
  • Versions
  • About
  • Compare
  • Similar models
  • FAQ

Phi 3.5 vision Specifications

Technical details for the model's default version.

Version
Phi-3.5-vision-instruct
Released
Aug 23, 2024
Knowledge cutoff
Unknown
Parameters
4.2B
Context window
N/A
Max output
N/A
Inputs
image, text
Outputs
text
Open weights
No
License
MIT

Phi 3.5 vision Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

No capability profile

No version of this model currently has enough benchmark coverage for a score.

Phi 3.5 vision Benchmark Results

Benchmark scores for Phi-3.5-vision-instruct.

9 rows
Columns

Show columns

POPE86.1%012100.0%CAug 11, 2026
ScienceQA91.3%011100.0%CAug 11, 2026
InterGPS36.3%0220.0%CAug 11, 2026
MMBench81.9%06937.5%CAug 11, 2026
TextVQA72.0%121521.4%CAug 11, 2026
ChartQA81.8%172430.4%CAug 11, 2026
AI2D78.1%30326.5%CAug 11, 2026
MathVista43.9%38392.6%CAug 11, 2026
MMMU43.0%61633.2%CAug 11, 2026

Phi 3.5 vision Arena Results

Preference and agent-evaluation results for the default version.

2 rows
Columns

Show columns

visionoverall144851.62,592N/AAug 6, 2026
vision style controloverall144920.82,592N/AAug 6, 2026

Phi 3.5 vision Runtime Performance

Provider-specific output speed and catalog latency for Phi-3.5-vision-instruct. Runtime does not affect the capability score.

No runtime data

No provider-specific speed or latency record is linked to the default version yet.

Browse runtime rankings

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

Phi 3.5 vision Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
N/A
Official provider
N/A
Lowest third-party
N/A
Tracked offerings
0
No provider prices

The default version has no current input or output token prices.

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

Phi 3.5 vision Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

1 rows
Columns

Show columns

Phi-3.5-vision-instructAug 23, 20240.04.2BN/AN/ANoMIT

What is Phi 3.5 vision?

Key information about Phi 3.5 vision and its available data.

2B-parameter open multimodal model with up to 128K context tokens. It emphasizes multi-frame image understanding and reasoning, boosting performance on single-image benchmarks while enabling multi-image comparison, summarization, and even video analysis.

The model underwent safety post-training for improved instruction-following, alignment, and robust handling of visual and text inputs, and is released under the MIT license.

Data as of 2026-08-11.

Phi 3.5 vision vs nearby models

Open a comparison with the three ranked models immediately above and below this model.

Phi 3.5 visionvsGemma 3n E2BPhi 3.5 visionvsMistral NeMoPhi 3.5 visionvsLlama 3.2 3BPhi 3.5 visionvsIBM Granite 4.0 TinyPhi 3.5 visionvsLlama 3.2 11BPhi 3.5 visionvsJamba 1.5 Mini

Models similar to Phi 3.5 vision

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#2560.0
MI

Phi 3.5 mini

Microsoft

0.0 LLMBoard

DetailsCompare
#2770.0
MI

Phi 4 Mini

Microsoft

0.0 LLMBoard

DetailsCompare
#233+4.9
MI

Phi 3.5 MoE

Microsoft

4.9 LLMBoard

DetailsCompare
#229+7.4
MI

Phi 4 multimodal

Microsoft

7.4 LLMBoard

DetailsCompare
#223+9.4
MI

Phi 4

Microsoft

9.4 LLMBoard

DetailsCompare
#211+13.8
MI

Phi 4 Mini Reasoning

Microsoft

13.8 LLMBoard

DetailsCompare

FAQ

Common questions about Phi 3.5 vision.

When was Phi 3.5 vision released?

Phi 3.5 vision's default version was released on Aug 23, 2024.

How much does Phi 3.5 vision cost?

No official standard PAYG price is currently available for Phi 3.5 vision.

Who created Phi 3.5 vision?

Phi 3.5 vision was created by Microsoft.

What is the context window for Phi 3.5 vision?

A context window is not available for the default version.

Is Phi 3.5 vision open weight?

No. The default version is not marked as having publicly available weights.

How many API providers offer Phi 3.5 vision?

No provider offering is currently linked to the default version.

What models should I compare Phi 3.5 vision with?

Nearby ranked alternatives include Gemma 3n E2B, Mistral NeMo, Llama 3.2 3B.