llmboard.ai
Benchmarks
CompareRankings
llmboard.ai
Benchmarks
CompareRankings
HomeModelsPhi 3.5 vision

Microsoft model product

Phi 3.5 vision

Phi-3.5-vision-instruct is a 4.2B-parameter open multimodal model with up to 128K context tokens. It emphasizes multi-frame image understanding and reasoning, boosting performance on single-image benchmarks while enabling multi-image comparison, summarization, and even video analysis. The model underwent safety post-training for improved instruction-following, alignment, and robust handling of visual and text inputs, and is released under the MIT license.

Updated Aug 10, 2026. Default version: Phi-3.5-vision-instruct

Compare
LLMBoard score0.0Phi-3.5-vision-instruct
Coverage80%6 benchmark families
Context windowN/ATokens
Official input priceN/AOfficial price unavailable

On this page

  • Specification
  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Versions
  • About
  • Compare
  • Similar models
  • FAQ

Model specification

Structured fields from the published default version.

Version
Phi-3.5-vision-instruct
Released
Aug 23, 2024
Knowledge cutoff
Unknown
Parameters
4.2B
Context window
N/A
Max output
N/A
Inputs
image, text
Outputs
text
Open weights
No
License
MIT

Capability profile

This profile uses the latest version under this unique model that has a calculated LLMBoard score. Arena and price are excluded.

Phi-3.5-vision-instruct category scores

Benchmark results

Published benchmark records for the scored version Phi-3.5-vision-instruct.

9 rows
Columns

Show columns

POPE0.912100.0%CAug 7, 2026
ScienceQA0.911100.0%CAug 7, 2026
InterGPS0.4220.0%CAug 7, 2026
MMBench0.86937.5%CAug 7, 2026
TextVQA0.7121521.4%CAug 7, 2026
ChartQA0.8172430.4%CAug 7, 2026
AI2D0.830326.5%CAug 7, 2026
MathVista0.438392.6%CAug 7, 2026
MMMU0.461633.2%CAug 7, 2026

Arena results

Preference and agent-evaluation signals from published Arena datasets.

4 rows
Columns

Show columns

visionoverall144851.62,592N/AAug 6, 2026
visionenglish144876.21,608N/AAug 6, 2026
vision style controloverall144920.82,592N/AAug 6, 2026
vision style controlenglish144936.31,608N/AAug 6, 2026

Pricing

Official vendor API PAYG pricing is summarized first. The table then lists individual provider offerings without treating their minimum as the official price.

Official API
N/A
Official provider
N/A
Lowest third-party
N/A
Tracked offerings
0
No published price snapshot

The default version has no provider offering with current input or output token prices.

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

Phi 3.5 vision versions

All published versions linked to this unique model. The score columns identify the version used by the current overall ranking.

1 rows
Columns

Show columns

Phi-3.5-vision-instructAug 23, 20240.04.2BN/AN/ANoMIT

What is Phi 3.5 vision?

A concise description based on the published model registry.

Phi-3.5-vision-instruct is a 4.2B-parameter open multimodal model with up to 128K context tokens. It emphasizes multi-frame image understanding and reasoning, boosting performance on single-image benchmarks while enabling multi-image comparison, summarization, and even video analysis. The model underwent safety post-training for improved instruction-following, alignment, and robust handling of visual and text inputs, and is released under the MIT license.

Use the benchmark, Arena and pricing sections above as separate evidence. A missing field means the current data snapshot does not support that claim.

Data snapshot: 2026-08-07. Editorial model content is not available in the backend.

Phi 3.5 vision vs nearby models

Open a comparison with the three ranked models immediately above and below this model.

Phi 3.5 visionvsGemma 3n E2BPhi 3.5 visionvsMistral NeMoPhi 3.5 visionvsLlama 3.2 3BPhi 3.5 visionvsIBM Granite 4.0 TinyPhi 3.5 visionvsLlama 3.2 11BPhi 3.5 visionvsJamba 1.5 Mini

Models similar to Phi 3.5 vision

Recommendations prioritize the same model type and family, then the closest published LLMBoard score.

#2560.0
MI

Phi 3.5 mini

Microsoft

0.0 LLMBoard

DetailsCompare
#2750.0
MI

Phi 4 Mini

Microsoft

0.0 LLMBoard

DetailsCompare
#235+5.1
MI

Phi 3.5 MoE

Microsoft

5.1 LLMBoard

DetailsCompare
#229+7.5
MI

Phi 4 multimodal

Microsoft

7.5 LLMBoard

DetailsCompare
#221+10.4
MI

Phi 4

Microsoft

10.4 LLMBoard

DetailsCompare
#211+14.0
MI

Phi 4 Mini Reasoning

Microsoft

14.0 LLMBoard

DetailsCompare

FAQ

Common questions about Phi 3.5 vision.

When was Phi 3.5 vision released?

Phi 3.5 vision's default version was released on Aug 23, 2024.

How much does Phi 3.5 vision cost?

No official standard PAYG price is currently available for Phi 3.5 vision.

Who created Phi 3.5 vision?

Phi 3.5 vision is published under Microsoft in the model registry.

What is the context window for Phi 3.5 vision?

The current registry does not publish a context window for the default version.

Is Phi 3.5 vision open weight?

No. The default version is not marked as having publicly available weights.

How many API providers offer Phi 3.5 vision?

No published provider offering is currently linked to the default version.

What models should I compare Phi 3.5 vision with?

Nearby ranked alternatives include Gemma 3n E2B, Mistral NeMo, Llama 3.2 3B.

Rankings

OverallCodingText ArenaPricing

Modalities

Image GenerationVideo GenerationSpeech-to-TextEmbeddings

Benchmarks

All BenchmarksReasoningMathCoding

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai