Embedding API pricing

Prices per 1M input tokens for 25 embedding models. "10K pages" assumes about 5M tokens (10,000 pages × ~400 words). Smaller vectors (fewer dimensions) also cut vector database storage costs.

ModelProviderPrice /1MDimensions10K pages
text-embedding-3-smallOpenAI$0.021536$0.10
voyage-3-liteVoyage AI$0.02—$0.10
voyage-3.5-liteVoyage AI$0.02—$0.10
voyage-4-liteVoyage AI$0.021024$0.10
voyage-3Voyage AI$0.06—$0.30
voyage-3.5Voyage AI$0.06—$0.30
voyage-4Voyage AI$0.061024$0.30
embed-v5.0-fastCohere$0.082048$0.40
embed-english-light-v3.0Cohere$0.10—$0.50
embed-english-v3.0Cohere$0.10—$0.50
embed-multilingual-v3.0Cohere$0.10—$0.50
mistral-embedMistral$0.10—$0.50
embed-v4.0Cohere$0.121536$0.60
embed-v5.0-proCohere$0.122048$0.60
voyage-4-largeVoyage AI$0.121024$0.60
voyage-code-4Voyage AI$0.121024$0.60
voyage-context-4Voyage AI$0.121024$0.60
voyage-multimodal-3Voyage AI$0.12—$0.60
voyage-multimodal-3.5Voyage AI$0.121024$0.60
text-embedding-3-largeOpenAI$0.133072$0.65
gemini-embedding-001Google$0.153072$0.75
codestral-embedMistral$0.15—$0.75
voyage-3-largeVoyage AI$0.18—$0.90
voyage-code-3Voyage AI$0.18—$0.90
voyage-context-3Voyage AI$0.18—$0.90

FAQ

How many tokens is my document set?

Roughly 1,300 tokens per 1,000 English words. 10,000 pages of about 400 words each is around 5M tokens. Use the token counter to check a sample.

Is embedding the expensive part of RAG?

Rarely. Embedding is a one-time cost per document (plus updates). The generation model reading retrieved chunks on every query usually dominates the monthly bill.