Inference host
Together AI
LLM prices per million tokens
together.ai
- Models
- 12
- Routes
- 12
- Cheapest
- $0.14
- DeepSeek V4 Flash 0731 input
- Batch
- Yes
- see note below
Models on Together AI
DeepSeek V4 Flash 0731Together AI$0.14 / $0.28+121% vs headline
Qwen3.5-9BTogether AI$0.17 / $0.25+105% vs headline
GLM 5.3 FlashTogether AI$0.15 / $0.5+245% vs headline
gpt-oss-120bTogether AI$0.15 / $0.6+304% vs headline
DeepSeek V4.1 FlashTogether AI$0.3 / $1.2+100% vs headline
Muse Glimmer 30BTogether AI$0.35 / $1.5+27% vs headline
Llama 3.3 70B InstructTogether AI$1.04 / $1.04+571% vs headline
DeepSeek V4 Pro 0813Together AI$1.32 / $3.96+100% vs headline
GLM 5.3Together AI$1.4 / $4.4+105% vs headline
GLM 5.2Together AI$1.4 / $4.4+147% vs headline
Qwen3.8 2.4T A95BTogether AI$2 / $6
Kimi K3Together AI$3 / $15+61% vs headline
Notes from the pricing page
Batch: Run asynchronous batch workloads at up to 50% lower cost (docs.together.ai batch-inference); only selected serverless models get 50% off
Caching: Cached input prices listed per model in tables (e.g. $0.30 input, $0.06 cached)
Cloud platform for training, fine-tuning and serving open-source AI models.
List prices from providers' pages on the date shown; your contract may differ. Logos via logo.dev; trademarks belong to their owners.