Skip to content
Inference host

Groq

LLM prices per million tokens

groq.com
Models
5
Routes
5
Cheapest
$0.05
Llama 3.1 8B Instruct input
Batch
Yes
see note below

Models on Groq

Notes from the pricing page

Batch: 50% cost discount compared to synchronous APIs (console.groq.com/docs/batch); does not stack with prompt caching

Caching: There is a 50% discount for cached input tokens (console.groq.com/docs/prompt-caching)

List prices from providers' pages on the date shown; your contract may differ. Logos via logo.dev; trademarks belong to their owners.

Weekly: LLM API list prices that changed, Thursday mornings.