OpenAI
Find the right AI model for your workload
Start with capabilities, pricing, context and model families—not a leaderboard alone.
Explore models by your constraints
Pricing stays connected to the existing llmprice.app data pipeline.
OpenAI
gpt-5.6-terra
OpenAI
gpt-5.6-luna
OpenAI
gpt-5.5
OpenAI
gpt-5.4
OpenAI
gpt-5.4-mini
OpenAI
gpt-5.4-nano
OpenAI
gpt-4o
OpenAI
gpt-4o-mini
OpenAI
o3
OpenAI
o4-mini
Anthropic
claude-opus-5
Anthropic
claude-fable-5
Anthropic
claude-mythos-5
Anthropic
claude-opus-4-8
Anthropic
claude-opus-4-6
Anthropic
claude-sonnet-5
Anthropic
claude-sonnet-4-6
Anthropic
claude-haiku-4-5
gemini-3.7-flash
gemini-3.6-flash
gemini-3.1-flash-lite
gemini-3.5-flash
gemini-3.1-pro
gemini-2.5-pro
gemini-2.5-flash
Mistral
mistral-large-3
Mistral
mistral-medium-3.5
Mistral
codestral-latest
DeepSeek
deepseek-v4-flash
DeepSeek
deepseek-chat-v4
DeepSeek
deepseek-reasoner-v4
xAI
grok-4.20
xAI
grok-4.5
xAI
grok-4.3
Together AI
meta-llama/Llama-3.3-70B-Instruct-Turbo
Together AI
Qwen/Qwen2.5-72B-Instruct-Turbo
Cohere
command-a
Cohere
command-r
Groq
llama-3.3-70b-versatile
Microsoft
Phi-4
Microsoft
Phi-4-mini
Microsoft
MAI-DS-R1
Microsoft
Llama-3.3-70B-Instruct
Zhipu
GLM-5.2
Zhipu
GLM-4.7
Alibaba
qwen3.8-max
Alibaba
qwen3.5
Alibaba
qwen3.5-plus
Alibaba
qwen3.5-flash
Alibaba
qwen-plus
Alibaba
qwen-turbo
Moonshot AI
kimi-k3
Moonshot AI
kimi-k2.6
ByteDance
doubao-seed-2.1-pro
ByteDance
doubao-seed-2.1-turbo
MiniMax
MiniMax-M3
Baidu
ERNIE 5.1
Baidu
ERNIE X1
Baidu
ERNIE 4.5
Tencent
Hunyuan HY3
Tencent
Hunyuan TurboS
Baichuan
Baichuan4
Baichuan
Baichuan4-Air
StepFun
Step 3.7 Flash
StepFun
Step 3.5 Flash
01.AI
Yi-Lightning
InclusionAI
Ling-3.0-flash
Choose by scenario, not leaderboard position
Start with coding, RAG or high-throughput needs, then validate price and specs.
New to model selection? Start here
01How should I compare API models?+
Define the quality bar first, then compare price, latency, context and modalities.
02Is a larger context window always better?+
No. More context can increase cost and latency; retrieval and caching also matter.
03How do I estimate monthly LLM cost?+
Multiply tokens per request by monthly volume, then apply input and output rates.
Model leaderboards — straight to the answer
Four rankings computed from the current dataset. Follow any model for full specs and monthly cost.
5 cheapest models
By blended rate at a 3:1 input-to-output mix
- 1.Phi-4-miniMicrosoft$0.025 / 1M
- 2.qwen3.5-flashAlibaba$0.055 / 1M
- 3.Phi-4Microsoft$0.088 / 1M
- 4.qwen-turboAlibaba$0.088 / 1M
- 5.Hunyuan HY3Tencent$0.10 / 1M
5 best-value models
Quality score ÷ output price
- 1.Phi-4-miniMicrosoft1700.0
- 2.qwen3.5-flashAlibaba630.8
- 3.qwen3.5Alibaba606.7
- 4.Yi-Lightning01.AI557.1
- 5.Phi-4Microsoft542.9
5 highest-quality models
Composite benchmark score
- 1.gpt-5.6-solOpenAIQuality 98
- 2.gpt-5.5OpenAIQuality 97
- 3.claude-opus-5AnthropicQuality 97
- 4.claude-fable-5AnthropicQuality 96
- 5.claude-mythos-5AnthropicQuality 96
5 largest context windows
For long documents and whole codebases
- 1.grok-4.20xAI2000K
- 2.gpt-5.6-solOpenAI1049K
- 3.gpt-5.6-terraOpenAI1049K
- 4.gpt-5.6-lunaOpenAI1049K
- 5.gpt-5.5OpenAI1049K