API pricing per million tokens
- Input
- $0.075
- per 1M, DeepInfra (fp8)
- Output
- $0.2
- per 1M tokens
- Cached input
- n/a
- no batch tier listed
- Context
- 256K
- 16K max output
Where to buy it (3)
DeepInfrafp8, fp8$0.075 / $0.2$0.106 blended
Venicefp8, fp8$0.094 / $0.25$0.133 blended
Parasailbf16, bf16$0.09 / $0.3cached $0.05
Through a gateway
Your own numbers100M input and 20M output tokens a month: $11.50 direct.
Helicone$11.50+0%
Kong AI Gateway$11.50+0%
LiteLLM$11.50+0%
Vercel AI Gateway$11.50+0%
AWS Bedrock$11.50+0%
Azure AI Foundry$11.50+0%
Google Vertex AI$11.50+0%
Databricks Mosaic AI$11.50+0%
Hugging Face Inference$11.50+0%
Cloudflare AI Gateway$12.07+5%
Requesty$12.07+5%
Eden AI$12.13+5.5%
OpenRouter$12.13+5.5%
Snowflake Cortex AI$13.80+20%
Martian Gatewayn/a
Portkeyn/a
TrueFoundry AI Gatewayn/a
IBM watsonx.ain/a
OCI Generative AIn/a
BigQuery MLn/a
Cloudflare Workers AIn/a
NVIDIA NIMn/a
Related models
Mistral Medium 3.5Mistral AI · 262K$1.5 / $7.5in / out per 1M$1.5$7.5n/a262KApr 30, 2026
Mistral Small 4Mistral AI · 262K$0.15 / $0.6in / out per 1M$0.15$0.6$0.015262KMar 16, 2026
Ministral 3 14B 2512Mistral AI · 262K$0.2 / $0.2in / out per 1M$0.2$0.2$0.02262KDec 2, 2025
Ministral 3 3B 2512Mistral AI · 131K$0.1 / $0.1in / out per 1M$0.1$0.1$0.01131KDec 2, 2025
GPT-6 LunaOpenAI · 1.05M$0.1 / $0.5in / out per 1M$0.1$0.5$0.011.05MSep 22, 2026
GPT-6 Luna ProOpenAI · 1.05M$0.1 / $0.5in / out per 1M$0.1$0.5$0.011.05MSep 22, 2026
Qwen3.8 Omni FlashQwen (Alibaba) · 1M$0.15 / $0.47in / out per 1M$0.15$0.47$0.0161MSep 21, 2026
Muse Spark 1.3 ContributorMeta · 1M$0.1 / $0.2in / out per 1M$0.1$0.2$0.0021MSep 2, 2026
About Mistral Small 3.2 24B
Mistral-Small-3.2-24B-Instruct-2506 is an updated 24B parameter model from Mistral optimized for instruction following, repetition reduction, and improved function calling.
Takes image, text in and returns text. First listed Sep 23, 2026. Blended price $0.106 per 1M tokens at 3 input to 1 output.
Questions
How much does Mistral Small 3.2 24B cost?
$0.075 per million input tokens and $0.2 per million output tokens.
What is the context window of Mistral Small 3.2 24B?
256K tokens, with up to 16K tokens of output.
Where is Mistral Small 3.2 24B cheapest?
Of 3 routes tracked, DeepInfra is cheapest at $0.075 input and $0.2 output per million tokens.
List prices from providers' pages on the date shown; your contract may differ. Logos via logo.dev; trademarks belong to their owners.