Skip to content
Inference host

Together AI

LLM prices per million tokens

together.ai
Official pricing
Models
12
Routes
12
Cheapest
$0.14
DeepSeek V4 Flash 0731 input
Batch
Yes
see note below

Models on Together AI

Notes from the pricing page

Batch: Run asynchronous batch workloads at up to 50% lower cost (docs.together.ai batch-inference); only selected serverless models get 50% off

Caching: Cached input prices listed per model in tables (e.g. $0.30 input, $0.06 cached)

Cloud platform for training, fine-tuning and serving open-source AI models.

List prices from providers' pages on the date shown; your contract may differ. Logos via logo.dev; trademarks belong to their owners.

Weekly: LLM API list prices that changed, Thursday mornings.