Skip to content

Nemotron 3 Embed 1BFree

NVIDIA Nemotron-3-Embed-1B is a 1.14B-parameter multilingual text embedding model (2048 dimensions, 34 languages) for semantic search, retrieval, and RAG. Built on Ministral-3-3B-Instruct; state-of-the-art on multilingual retrieval benchmarks (MMTEB 71.05, RTEB 72.38). Served via NVIDIA NIM.

Providers
Capabilities
NVIDIA
nvidia
$0
$0
NVIDIA
nvidia-byok
$0
$0
Unavailable
Usage analytics

Loading usage…

Uptime & Health
No uptime data yet

These providers haven't been health-probed for this model yet. The router still routes around upstreams that fail live requests — uptime fills in once probe history accrues.

Share cards
Nemotron 3 Embed 1B share card
Nemotron 3 Embed 1B
NVIDIA upstream share card
NVIDIA upstream
NVIDIA upstream share card
NVIDIA upstream
Credits
Use your own key

Run Nemotron 3 Embed 1B on your own key — your requests are billed by the provider. Pool callers pay AnyRouter credits.

No BYOK keys configured for this model yet.

Share a key with the pool to earn credits for every request it serves, covering your plan cost.

Embedding vectors
Vector dimensionsNot published
Max input32,768 tokens
PriceFree on AnyRouter
Request parameters
inputmodelencoding_format
ArchitectureTransformer
Categoryembedding
ReleasedJul 16, 2026
Modalities
→
Capabilities
Embeddings are fixed-length vectors — compare them with cosine similarity for semantic search, RAG retrieval, clustering, and deduplication. Embed queries and documents with the same model, or the distances are meaningless.