Inference host
Groq
LLM prices per million tokens
groq.com
- Models
- 5
- Routes
- 5
- Cheapest
- $0.05
- Llama 3.1 8B Instruct input
- Batch
- Yes
- see note below
Models on Groq
Notes from the pricing page
Batch: 50% cost discount compared to synchronous APIs (console.groq.com/docs/batch); does not stack with prompt caching
Caching: There is a 50% discount for cached input tokens (console.groq.com/docs/prompt-caching)
List prices from providers' pages on the date shown; your contract may differ. Logos via logo.dev; trademarks belong to their owners.