Microsoft model product
8B parameters) open model built upon synthetic data and filtered web data, focusing on high-quality reasoning.
Updated Aug 12, 2026. Default version: Phi 4 Mini
Technical details for the model's default version.
This profile uses the model's current scored version. Arena ratings and prices are shown separately.
Benchmark scores for Phi 4 Mini.
| OpenBookQA | 0.8 | 3 | 5 | 50.0% | C | |
| Social IQa | 0.7 | 3 | 9 | 75.0% | C | |
| TruthfulQA | 0.7 | 4 | 18 | 82.3% | C | |
| Multilingual MMLU | 0.5 | 5 | 5 | 0.0% | C | |
| BoolQ | 0.8 | 7 | 10 | 33.3% | C | |
| PIQA | 0.8 | 10 | 11 | 10.0% | C | |
| BIG-Bench Hard | 0.7 | 12 | 21 | 45.0% | C | |
| ARC-C | 0.8 | 15 | 34 | 57.6% | C | |
| Winogrande | 0.7 | 19 | 22 | 14.3% | C | |
| Arena Hard | 0.3 | 24 | 26 | 8.0% | C | |
| MGSM | 0.6 | 24 | 31 | 23.3% | C | |
| HellaSwag | 0.7 | 26 | 27 | 3.9% | C | |
| GSM8k | 0.9 | 32 | 48 | 34.0% | C | |
| MATH | 0.6 | 45 | 71 | 37.1% | C | |
| MMLU | 0.7 | 90 | 100 | 10.1% | C | |
| MMLU-Pro | 0.5 | 111 | 129 | 14.1% | C | |
| GPQA | 0.3 | 228 | 234 | 2.6% | C |
Preference and agent-evaluation results for the default version.
The default version does not have a matching Arena result yet.
Official vendor API pricing appears first, followed by individual provider offers.
| Nvidia | microsoft/phi-4-mini-instruct | global | N/A | N/A | 131.1K | |
| Azure Cognitive Services | phi-4-mini | global | $0.075 | $0.30 | 128K | |
| Azure | phi-4-mini | global | $0.075 | $0.30 | 128K | |
| NanoGPT | phi-4-mini-instruct | global | $0.17 | $0.68 | 128K |
Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.
Available versions of this model. The score column identifies the version used in the overall ranking.
| Phi 4 Mini | 0.0 | 3.8B | 128K | 4.1K | Yes | MIT |
Key information about Phi 4 Mini and its available data.
8B parameters) open model built upon synthetic data and filtered web data, focusing on high-quality reasoning. It supports a 128K token context length and is enhanced for instruction adherence and safety via supervised fine-tuning and direct preference optimization.
Data as of 2026-08-11.
Open a comparison with the three ranked models immediately above and below this model.
Recommendations prioritize the same model type and family, then the closest LLMBoard score.
Common questions about Phi 4 Mini.
Phi 4 Mini's default version was released on Feb 1, 2025.
Phi 4 Mini's official API price is $0.075 per million input tokens and $0.30 per million output tokens via Azure. The lowest tracked third-party offer starts at $0.075 input and $0.30 output via Azure Cognitive Services.
Phi 4 Mini was created by Microsoft.
The default version has a 128K token context window.
Yes. The default version is marked as open weight under MIT.
4 provider offerings are linked to the default version.
Nearby ranked alternatives include Llama 3.1 8B, Qwen2 7B, ERNIE 4.5.