LLM Insights中文
Groq logo

Groq API Pricing & Model Comparison

We track 1 Groq models, with input pricing from $0.59 to $0.59 per million tokens. The line-up's blended rate runs 77% below the catalog average. Below: price, context, speed and value compared across the range.

Models tracked1
Lowest input price$0.59 / 1M
Top quality score78
Largest context128K

For cost, Groq's cheapest option is llama-3.3-70b-versatile ($0.59 input / $0.79 output); the quality ceiling is llama-3.3-70b-versatile at a quality score of 78. On a 50,000-turn-per-month support workload that is $79 versus $79 — the spread inside a single provider is often wider than the spread between providers.

On value (quality score ÷ output price) the pick of the Groq range is llama-3.3-70b-versatile. The largest context window is 128K, offered only by llama-3.3-70b-versatile, and the fastest output is llama-3.3-70b-versatile at 300 tok/s. Modalities across the line-up: Text.

Models and pricing

Sorted by blended rate, cheapest first (3:1 input-to-output mix). Follow a model name for full specs and monthly cost estimates.

ModelInput /1MOutput /1MBlended /1MQualityContextSpeedValueUse cases
llama-3.3-70b-versatile$0.59$0.79$0.6478128K300 tok/s98.7
ChatbotAffordableFast

Frequently asked questions

Which Groq model is cheapest?

llama-3.3-70b-versatile, at $0.59 input and $0.79 output per million tokens, with a quality score of 78.

What is Groq's flagship model?

By quality score it is llama-3.3-70b-versatile (78), priced at $0.59 input / $0.79 output with a 128K context window.

Does Groq pricing change?

Yes. llmprice.app runs a daily collection job and refreshes these figures, but there can be a lag after a provider re-prices — confirm against Groq's official pricing page before committing.

Browse all providers

Prices and specs are collected daily by llmprice.app. Confirm against the official pricing page before production use.