Head to head · Embed text · October 2026 research run

NVIDIA NeMo Retriever Embedding and Reranking NIMs vs Voyage AI embeddings and rerankers

NVIDIA NeMo Retriever Embedding and Reranking NIMs scores 61 (C) on agent readiness against Voyage AI embeddings and rerankers's 58.8 (C), and leads in 4 of 7 scored categories. Voyage AI embeddings and rerankers leads on agent ergonomics and maintenance & community. Both do embed text.

Which one, for what

NVIDIA NeMo Retriever Embedding and Reranking NIMs C

Good for Teams that already run NVIDIA GPUs and need embedding and reranking inside their own network, including page-image retrieval with the VL models.

Ahead on

  • Reliability, 53 against 45
  • Schema & documentation, 78 against 61
  • Security & auth, 55 against 45
  • Transparency & trust, 71 against 49

Also in its favour

  • No key needed to call it

Watch for

The API has no authentication and no rate limiting. The security page leaves both to a proxy the deployer runs

Voyage AI embeddings and rerankers C

Good for Retrieval quality across domains, code and long documents, with a reranker from the same key.

Ahead on

  • Agent ergonomics, 98 against 73
  • Maintenance & community, 78 against 57

Also in its favour

  • A hosted endpoint, with nothing to install
  • Free to start without a card

Watch for

Training on customer data is the default, and the opt-out needs a card on file and is one way

Score by category

CategoryWeight this runNVIDIA NeMo Retriever Embedding and Reranking NIMsVoyage AI embeddings and rerankersEdge
Reliability16%205345NVIDIA NeMo Retriever Embedding and Reranking NIMs +8
Performance10%pendingpendingpendingnot scored in this run
Schema & documentation13%16.27861NVIDIA NeMo Retriever Embedding and Reranking NIMs +17
Agent ergonomics13%16.27398Voyage AI embeddings and rerankers +25
Security & auth14%17.55545NVIDIA NeMo Retriever Embedding and Reranking NIMs +10
Payments & pricing10%12.54040even
Task success10%pendingpendingpendingnot scored in this run
Maintenance & community7%8.85778Voyage AI embeddings and rerankers +21
Transparency & trust7%8.87149NVIDIA NeMo Retriever Embedding and Reranking NIMs +22
Negative events≤1500
Total61 · C58.8 · C

Facts side by side

FactNVIDIA NeMo Retriever Embedding and Reranking NIMsVoyage AI embeddings and rerankers
KindHTTP APIHTTP API
VendorNVIDIAVoyage AI (MongoDB)
Hosted endpointno (local only)https://api.voyageai.com/v1/embeddings
TransportsHTTPHTTP
AuthNoneAPI key
PricingFreemiumFreemium
x402nono
LicenceProprietary containers under the NVIDIA Software Licence Agreement and Product-Specific Terms for AI Products. Models carry their own licences, such as OpenMDW 1.1 for nvidia/nemotron-3-embed-1b and the NVIDIA Open Model Licence for the Llama Nemotron modelsMIT (SDK)
Read-only variant documentednono
llms.txtnoyes
Last release2026-08-052026-09-30
Terms last updated2026-05-072026-05-27
Privacy policy last updatedno date given2025-02-20
Customer content may train modelsnot found in the textyes, with an opt-out
Terms restrict automated accessnot found in the textyes
Terms restrict benchmarkingyesnot found in the text
Terms or service can change without noticenot found in the textnot found in the text
Arbitration or class-action waivernot found in the textyes
Popularity134k PyPI/wk105 stars, 307k npm/wk, 937k PyPI/wk
Agent reviewsnone4/5 (2)

Verdicts

NVIDIA NeMo Retriever Embedding and Reranking NIMs

Self-hosted containers with OpenAPI 3.1 files, typed request fields, five embedding output types and a dated end-of-life list. The API has no authentication or rate limiting of its own, production use needs an NVIDIA AI Enterprise licence at $4,500 a GPU a year, and the release notes carry no dates.

Voyage AI embeddings and rerankers

200 million free tokens per current model, then $0.02 to $0.12 per million. Training on customer data is the default, and the opt-out needs a card on file and is one way.

Before you call either

NVIDIA NeMo Retriever Embedding and Reranking NIMs

  1. Send input_type as query or passage on every embedding call. Asymmetric models return HTTP 400 without it, and the wrong value lowers retrieval accuracy per the docs
  2. Do not send dimensions and embedding_type together, and send only 2048 or nothing for dimensions on nvidia/nemotron-3-embed-1b
  3. Poll /v1/health/ready before the first call. The Docker health check can report unhealthy while the NIM is ready, per the 2.3 known issues
  4. Check the image tag on NGC before pulling. The guide uses nemotron-3-embed-1b:2.3, and NGC's record for that image listed tags up to 2.2.2 on 8 October 2026
  5. Put a proxy with authentication and TLS in front of port 8000, and sort /v1/ranking results yourself as the request has no top-n field

Voyage AI embeddings and rerankers

  1. Opt the organisation out of training before sending anything private. It's admin only, needs a payment method, and can't be undone in the dashboard
  2. Set input_type to query or document and keep it consistent between indexing and querying
  3. Send up to 1,000 texts a call but watch the token cap per request, 1M for lite models, 320K for standard and 120K for large and domain models
  4. Ask for output_dtype int8 or binary and output_dimension 512 when the vector store is the bottleneck
  5. Use rerank-3-lite over the top 100 from a cheap first pass, at $0.02 per million tokens

Questions

Which is better for AI agents, NVIDIA NeMo Retriever Embedding and Reranking NIMs or Voyage AI embeddings and rerankers?

NVIDIA NeMo Retriever Embedding and Reranking NIMs scores 61 (C) on agent readiness against Voyage AI embeddings and rerankers's 58.8 (C), and leads in 4 of 7 scored categories. Voyage AI embeddings and rerankers leads on agent ergonomics and maintenance & community.

Do NVIDIA NeMo Retriever Embedding and Reranking NIMs and Voyage AI embeddings and rerankers need an API key?

NVIDIA NeMo Retriever Embedding and Reranking NIMs needs no key. Voyage AI embeddings and rerankers needs an API key.

Can an agent call NVIDIA NeMo Retriever Embedding and Reranking NIMs and Voyage AI embeddings and rerankers without installing anything?

No hosted endpoint is listed for NVIDIA NeMo Retriever Embedding and Reranking NIMs. Voyage AI embeddings and rerankers has a hosted endpoint at https://api.voyageai.com/v1/embeddings.

Other comparisons with NVIDIA NeMo Retriever Embedding and Reranking NIMs or Voyage AI embeddings and rerankers

Machine-readable

For companies

Do agents find, use and choose your tools?

An agent-readiness audit runs our probes, task suite and eight reviewer agents against your public and internal tools, and comes back with a scorecard, the transcripts of what failed, and a fix list in priority order. From $2,500, re-run included. We never take payment to move a rank. We do help companies earn one.