Skip to content

Qwen3 Embedding 8B

Qwen (Alibaba) · Open weights · 33K context · released Oct 28, 2025 · checked 9h ago

Ways to buy 3

Way to buyIn / out per 1M
DeepInfra
us
$0.01 / FreeCheapest
Nebius
Default region
$0.01 / FreeCheapest
SiliconFlow
fp8
$0.04 / Free+300% vs cheapest

USD per 1M tokens. Regions are separate routes.

Through a gateway

Your own numbers

100M input and 20M output tokens a month: $1.00 direct.

No published token fee: Martian Gateway, Portkey, TrueFoundry AI Gateway.

Related models

In / out per 1M
Qwen3.8 Max Prime
Qwen (Alibaba) · 1M
$4 / $12
Qwen3.8 Omni Flash
Qwen (Alibaba) · 1M
$0.15 / $0.47
Qwen3.8 Max (0902)
Qwen (Alibaba) · 1M
$2 / $6
Qwen3.8 Flash
Qwen (Alibaba) · 1M
$0.15 / $0.47
Text Embedding 3 Small
OpenAI · 8K
$0.02 / Free

About Qwen3 Embedding 8B

The Qwen3 Embedding model series is the latest proprietary model of the Qwen family, specifically designed for text embedding and ranking tasks.

Takes text in and returns embeddings. First listed Sep 24, 2026. Up to 32K output tokens. Blended price $0.0075 per 1M tokens at 3 input to 1 output. Model id qwen3-embedding-8b.

Price history

No list-price history: this model is sold by hosts rather than its lab.

Across fru.devCompany profile

Questions

How much does Qwen3 Embedding 8B cost?

$0.01 per million input tokens and Free per million output tokens.

What is the context window of Qwen3 Embedding 8B?

33K tokens, with up to 32K tokens of output.

Where is Qwen3 Embedding 8B cheapest?

Of 3 routes tracked, DeepInfra is cheapest at $0.01 input and Free output per million tokens.

List prices from providers' pages on the date shown; your contract may differ. Logos via logo.dev; trademarks belong to their owners.

Weekly: LLM API list prices that changed, Thursday mornings.