verifier.org

voyage-4-large

cheaper than OpenAI, four times the context, and a reranker to match

$0.12 / 1m tokens#retrieval-focused-teams-#reranking-pair-from-one-

beats the default on price, context and reranking, from a smaller vendor that publishes less about how it performs.

voyage-4-large is the current flagship, having replaced voyage-3-large during the period we were researching this category — and it cut the price on the way, from $0.18 per million to $0.12. that undercuts OpenAI while offering 32,000 tokens of context against OpenAI's 8,192.

dimensions default to 1,024 with 256, 512 and 2,048 also selectable, and Voyage states that all embeddings created with the 4 series are compatible across those options, which is a more useful guarantee than it sounds — it means changing your mind about index size later does not force a full re-embed the way moving between Gemini generations does.

the description above is ours, condensed from the ranking. pricing moves — check it on the vendor's own page before you rely on it.

pricing
$0.12 / 1m tokens
our verdict

beats the default on price, context and reranking, from a smaller vendor that publishes less about how it performs.

we researched this category against vendors' own pricing pages and licence files. that is where this line comes from — not from the vendor, and not from anything they paid for.

more embedding models

text-embedding-3-small

openai's cheaper embedding model, shortenable to fewer dimensions

#api#matryoshka

multilingual-e5

microsoft's open multilingual embedding family, widely used as a baseline

#open-weights#multilingual

ColBERT

late-interaction retrieval that scores token by token rather than one vector

#late-interaction#research

Qwen3-Embedding-8B

we checked this$0.010 / 1m tokens on DeepInfra

apache-2.0, best open multilingual quality, and a tenth of a cent hosted

#multilingual-retrieval-w#the-bill-has-to-be-small