Groq API Pricing & Model Comparison
We track 1 Groq models, with input pricing from $0.59 to $0.59 per million tokens. The line-up's blended rate runs 77% below the catalog average. Below: price, context, speed and value compared across the range.
For cost, Groq's cheapest option is llama-3.3-70b-versatile ($0.59 input / $0.79 output); the quality ceiling is llama-3.3-70b-versatile at a quality score of 78. On a 50,000-turn-per-month support workload that is $79 versus $79 — the spread inside a single provider is often wider than the spread between providers.
On value (quality score ÷ output price) the pick of the Groq range is llama-3.3-70b-versatile. The largest context window is 128K, offered only by llama-3.3-70b-versatile, and the fastest output is llama-3.3-70b-versatile at 300 tok/s. Modalities across the line-up: Text.
Models and pricing
Sorted by blended rate, cheapest first (3:1 input-to-output mix). Follow a model name for full specs and monthly cost estimates.
| Model | Input /1M | Output /1M | Blended /1M | Quality | Context | Speed | Value | Use cases |
|---|---|---|---|---|---|---|---|---|
| llama-3.3-70b-versatile | $0.59 | $0.79 | $0.64 | 78 | 128K | 300 tok/s | 98.7 | ChatbotAffordableFast |
Frequently asked questions
Which Groq model is cheapest?
llama-3.3-70b-versatile, at $0.59 input and $0.79 output per million tokens, with a quality score of 78.
What is Groq's flagship model?
By quality score it is llama-3.3-70b-versatile (78), priced at $0.59 input / $0.79 output with a 128K context window.
Does Groq pricing change?
Yes. llmprice.app runs a daily collection job and refreshes these figures, but there can be a lag after a provider re-prices — confirm against Groq's official pricing page before committing.
Prices and specs are collected daily by llmprice.app. Confirm against the official pricing page before production use.