Voyage AI / voyageai/rerank-3-lite

Voyage Rerank 3 Lite - access through LLMTR

Voyage Rerank 3 Lite uses the same request shape and the same limits as Rerank 3 at two and a half times lower cost per million tokens. Voyage offers it as the fast, economical option optimized for latency-sensitive applications. It suits search boxes where every user query is reranked, in-chat retrieval and production workloads where volume drives cost. The query can be at most 8,000 tokens and the query plus any single document at most 32,000 tokens; a single request processes at most 1000 documents and 600,000 total tokens.

Technical specifications

Canonical IDvoyageai/rerank-3-lite
ProviderVoyage AI
Context window32,000 tokens
OperationsRERANK
Modalitiestext

Pricing

An 8% platform margin applies to credit top-ups; model usage prices are not separately marked up.

OperationMetricUnitPrice
RERANKINPUT_TEXTPER_1M_TOKENS$0.020000

Example usage

With existing OpenAI SDK flows, change only the base URL and model identifier.

curl https://llmtr.com/v1/chat/completions   -H "Authorization: Bearer llmtr-your_key"   -H "Content-Type: application/json"   -d '{"model":"voyageai/rerank-3-lite","messages":[{"role":"user","content":"Hello"}]}'

Guides about this model

Related models