Microsoft API Pricing & Model Comparison
We track 5 Microsoft models, with input pricing from $0.02 to $2.00 per million tokens. The line-up's blended rate runs 75% below the catalog average. Below: price, context, speed and value compared across the range.
For cost, Microsoft's cheapest option is Phi-4-mini ($0.02 input / $0.04 output); the quality ceiling is MAI-DS-R1 at a quality score of 90. On a 50,000-turn-per-month support workload that is $3.00 versus $21 — the spread inside a single provider is often wider than the spread between providers.
On value (quality score ÷ output price) the pick of the Microsoft range is Phi-4-mini. The largest context window is 128K, shared by 4 models in the range, and the fastest output is Phi-4-mini at 250 tok/s. Modalities across the line-up: Text, Vision.
Models and pricing
Sorted by blended rate, cheapest first (3:1 input-to-output mix). Follow a model name for full specs and monthly cost estimates.
| Model | Input /1M | Output /1M | Blended /1M | Quality | Context | Speed | Value | Use cases |
|---|---|---|---|---|---|---|---|---|
| Phi-4-mini | $0.02 | $0.04 | $0.025 | 68 | 128K | 250 tok/s | 1700.0 | ChatbotAffordableFast |
| Phi-4 | $0.07 | $0.14 | $0.088 | 76 | 16K | 180 tok/s | 542.9 | CodingChatbotAffordableFast |
| Llama-3.3-70B-Instruct | $0.10 | $0.32 | $0.16 | 78 | 128K | 100 tok/s | 243.8 | ChatbotCoding |
| MAI-DS-R1 | $0.14 | $0.28 | $0.18 | 90 | 128K | 60 tok/s | 321.4 | ReasoningCoding |
| Mistral-Large | $2.00 | $6.00 | $3.00 | 82 | 128K | 80 tok/s | 13.7 | ChatbotCoding |
Frequently asked questions
Which Microsoft model is cheapest?
Phi-4-mini, at $0.02 input and $0.04 output per million tokens, with a quality score of 68.
What is Microsoft's flagship model?
By quality score it is MAI-DS-R1 (90), priced at $0.14 input / $0.28 output with a 128K context window.
Does Microsoft pricing change?
Yes. llmprice.app runs a daily collection job and refreshes these figures, but there can be a lag after a provider re-prices — confirm against Microsoft's official pricing page before committing.
Prices and specs are collected daily by llmprice.app. Confirm against the official pricing page before production use.