Qwen3 Embedding 8B

Qwen3-based text embedding model with strong multilingual retrieval performance. Output dimension 4096; supports the `dimensions` parameter to shorten output.

qwen3-embedding-8b
Embedding modelSTABLEGet StartedView uptime
40,960 context
Starting at $0.01/M input tokens
Starting at $0.00/M output tokens
Embeddings
No ratings yetSign in to rate

Select Provider

All Providers for Qwen3 Embedding 8B

LLM Gateway routes requests to the best providers that are able to handle your prompt size and parameters.

Nebius AI
Context: 41.0k
Input
$0.01
/M tokens
Cache Read
/M tokens
Output
$0
/M tokens
Get Started