Skip to content

Qwen3 Embedding 8B

Qwen3 8B embedding model for text embedding tasks with a 32,768 token context window.

Providers
Capabilities
DeepInfra
deepinfra-byok
$0
$0
Unavailable
OVHcloud
ovhcloud-byok
$0
$0
Unavailable
Usage analytics

Loading usage…

Uptime & Health
No uptime data yet

These providers haven't been health-probed for this model yet. The router still routes around upstreams that fail live requests — uptime fills in once probe history accrues.

Share cards
Qwen3 Embedding 8B share card
Qwen3 Embedding 8B
DeepInfra upstream share card
DeepInfra upstream
OVHcloud upstream share card
OVHcloud upstream
Credits
Use your own key

Run Qwen3 Embedding 8B on your own key — your requests are billed by the provider. Pool callers pay AnyRouter credits.

No BYOK keys configured for this model yet.

Share a key with the pool to earn credits for every request it serves, covering your plan cost.

Embedding vectors
Vector dimensions4,096
Max input32,768 tokens
Price$0.01 / 1M tokens
Request parameters
inputmodeldimensionsencoding_format
ArchitectureTransformer
Categoryembedding
ReleasedJun 4, 2025
Modalities
→
Capabilities
Embeddings are fixed-length vectors — compare them with cosine similarity for semantic search, RAG retrieval, clustering, and deduplication. Embed queries and documents with the same model, or the distances are meaningless.