Embedding Model Comparison: OpenAI vs Cohere vs Voyage vs Google vs Self-Hosted (June 2026)
Side-by-side comparison of every major embedding model. Price, dimensions, MTEB score, context window, batch discount, free tier, and best-fit use case.
Full Model Comparison Table
| Model | Provider | $/M std | $/M batch | Dims | Context | MTEB | Free tier | Best for |
|---|---|---|---|---|---|---|---|---|
| text-embedding-3-small | OpenAI | $0.020 | $0.010 | 1,536 | 8,191 | 62.3 | - | General-purpose RAG, low cost |
| text-embedding-3-large | OpenAI | $0.130 | $0.065 | 3,072 | 8,191 | 64.6 | - | High-accuracy retrieval |
| text-embedding-ada-002 | OpenAI | $0.100 | No batch | 1,536 | 8,191 | 60.5 | - | Legacy - migrate to 3-small |
| embed-v4 | Cohere | $0.120 | No batch | 1,536 | 128,000 | 55 | 100 calls/min | Multilingual content (100+ languages) |
| voyage-4-large | Voyage AI | $0.120 | $0.080 | 1,024 | 32,000 | - | 200M tokens | State-of-the-art retrieval (MoE) |
| voyage-4 | Voyage AI | $0.060 | $0.040 | 1,024 | 32,000 | - | 200M tokens | Best price-to-accuracy ratio |
| voyage-4-lite | Voyage AI | $0.020 | $0.013 | 1,024 | 32,000 | - | 200M tokens | High-volume, cost-sensitive |
| gemini-embedding-2-preview | $0.200 | No batch | 3,072 | 8,192 | 68 | - | Multimodal, Google ecosystem | |
| gemini-embedding-001 | $0.150 | No batch | 3,072 | 2,048 | 65.4 | - | Stable GA, Google ecosystem | |
| Titan Text Embeddings V2 | Amazon Bedrock | $0.020 | No batch | 1,024 | 8,192 | 62.8 | - | AWS-native apps, compliance |
| BGE-M3 (self-hosted) | Self-Hosted | $0.001 | $0.001 | 1,024 | 8,192 | 66.5 | - | High-volume, privacy-sensitive |
Green = best in category, amber = watch out, purple = top MTEB. MTEB Retrieval average where publicly available. Context window in tokens. Verified June 2026.
By Scenario: Which Model Should You Pick?
Under $5/month total. One-time embedding cost for 50M tokens is $1.00. pgvector is near-zero marginal cost on existing Postgres.
Provider detailsApproaches voyage-3-large quality at a mid-sized model's cost. At $0.06/M standard or $0.04/M batch, the accuracy premium is affordable for most production apps, with 200M free tokens to start.
Provider details15-20% quality improvement for non-Latin scripts over OpenAI. Its 128K-token context handles long documents with minimal chunking.
Provider detailsTops Voyage's RTEB benchmark (+14% vs OpenAI 3-large on NDCG@10) with a mixture-of-experts design. At $0.12/M, index with large and query with voyage-4-lite ($0.02/M) using the shared-vector-space feature.
Provider detailsKeeps data within AWS VPC. Binary embedding option for 32x storage reduction. Integrates with OpenSearch Serverless and SageMaker.
Provider detailsAt A100 spot rates, $0.001/M tokens vs $0.02/M for OpenAI small. Breaks even at roughly 15M tokens/month. Requires DevOps investment.
Provider detailsPurpose-trained on code repositories at $0.18/M. Meaningfully better than general-purpose models for code retrieval, with 200M free tokens to evaluate it.
Provider detailsStable GA model, $0.15/M, strong MTEB at 65.4. MRL support for dimension reduction. Integrates naturally with Vertex AI, BigQuery, and Cloud Storage.
Provider details