Skip to content

Qwen3 Embedding 4B

Qwen3 Embedding 4B is the 4B size in the Qwen3 embedding and reranking family, between the 0.6B and 8B cuts. It embeds text over a 32,768-token window at 2,560 dimensions, supports 100+ languages, is instruction-aware for task prefixes, and supports Matryoshka truncation from 32 to 2,560 dimensions. It ranks below the 8B cut on MTEB multilingual but well above the 0.6B cut, at a much lower cost.

Providers
Capabilities
OpenRouter
openrouter-byok
$0
$0
Unavailable
Usage analytics

Loading usage…

Uptime & Health
No uptime data yet

These providers haven't been health-probed for this model yet. The router still routes around upstreams that fail live requests — uptime fills in once probe history accrues.

Share cards
Qwen3 Embedding 4B share card
Qwen3 Embedding 4B
OpenRouter upstream share card
OpenRouter upstream
Credits
Use your own key

Run Qwen3 Embedding 4B on your own key — your requests are billed by the provider. Pool callers pay AnyRouter credits.

No BYOK keys configured for this model yet.

Share a key with the pool to earn credits for every request it serves, covering your plan cost.

Embedding vectors
Vector dimensions2,560
Max input32,768 tokens
Price$0.02 / 1M tokens
Request parameters
inputmodeldimensionsencoding_format
ArchitectureTransformer
Categoryembedding
ReleasedOct 28, 2025
Modalities
→
Capabilities
Embeddings are fixed-length vectors — compare them with cosine similarity for semantic search, RAG retrieval, clustering, and deduplication. Embed queries and documents with the same model, or the distances are meaningless.