Cohere Embed 4
Cohere's enterprise embedding model for multilingual text, images, and visually rich documents with a 128K context window.
Cohere's enterprise embedding model for multilingual text, images, and visually rich documents with a 128K context window.
Plain-English overview
Cohere Embed 4 is designed for search over normal prose and documents whose meaning depends on layout or visuals. It can embed text, images, and mixed-content PDFs, which is useful for slide decks, reports, diagrams, and scanned business material.
The endpoint supports several vector sizes and compact numerical formats. Query, document, classification, and clustering input types let the same model shape representations around the task rather than treating every input identically.
Category comparison
These are provider-published specifications, not Cody benchmark scores. Follow the linked sources for current limits and endpoint-specific exceptions.
Results describe a specific test, language and configuration—not overall intelligence. Missing results do not imply worse quality.
Community-submitted result
Retrieval relevance (0–1; higher is better)
MTEB · 1.38.43 · eng-Latn · test/default
dimensions: 1536 · similarity: cosine · modelMetadata: https://github.com/embeddings-benchmark/results/blob/main/results/Cohere__Cohere-embed-v4.0/1/model_meta.json
Checked: September 5, 2026
MTEB contributors · FinanceBenchRetrieval0.8833 nDCG@10
Pricing & comparisons
Set your usage. Your estimate updates as you type.
Assumes 600 tokens per page, processed separately. Actual token counts vary. This covers embedding only, not storage, search, or generated answers.
Cohere Embed 4
Cohere
Estimated total (USD)
For the usage above · USD · API pricing, not a subscription
Estimates exclude taxes, tools, cache storage/writes, free allowances and custom discounts. Image estimates cover output only, not prompt or reference-image charges. Quality modes differ by model. Unlisted settings are not treated as free.
API and provider access
Availability
Available through Cohere's Embed API, with private deployment and selected cloud-platform routes for enterprise customers.
Processing location and residency depend on the Cohere platform, private environment, or chosen cloud marketplace agreement.
Check live availabilityData and training
Cohere enterprise customers can opt out of training; SaaS prompts and generations are generally deleted after 30 days. Approved zero-data-retention accounts and private deployments offer stronger controls.
This is a concise reading of the cited provider material, not legal advice. A third-party gateway can have different storage, routing, training, and residency terms from the model maker's direct API.
Read the provider policyFrequently asked questions
Cohere's enterprise embedding model for multilingual text, images, and visually rich documents with a 128K context window. Cohere Embed 4 is designed for search over normal prose and documents whose meaning depends on layout or visuals. It can embed text, images, and mixed-content PDFs, which is useful for slide decks, reports, diagrams, and scanned business material.
Cohere Embed 4 was released on April 15, 2025 according to the cited provider materials.
Available through Cohere's Embed API, with private deployment and selected cloud-platform routes for enterprise customers. The access routes listed in this guide are Cohere, Microsoft Foundry, and AWS SageMaker.
Hosted unit price not published. Cohere's reviewed public pricing page does not state one comparable hosted Embed 4 token rate. Model Vault and private deployments use capacity or contract pricing.
Available through Cohere's Embed API, with private deployment and selected cloud-platform routes for enterprise customers. Processing location and residency depend on the Cohere platform, private environment, or chosen cloud marketplace agreement.
Cohere enterprise customers can opt out of training; SaaS prompts and generations are generally deleted after 30 days. Approved zero-data-retention accounts and private deployments offer stronger controls. The policy belongs to the provider route and account terms, so verify it again before production use.
USD per million tokens for the shown API routes at standard context length. Cache and long-context rates may differ. Models without matching reviewed prices are omitted; lower cost does not mean better quality.
Mistral AI
Voyage AI
Voyage AI
Voyage AI
OpenAI
Related comparisons
Google's multimodal embedding model for placing text, images, video, audio, and PDFs in one searchable vector space.
Open full comparisonOpenAI's highest-capability text embedding model for semantic search, recommendations, clustering, and retrieval pipelines.
Open full comparisonVoyage AI's quality-first general embedding model for text and code retrieval, with adjustable dimensions and a shared family vector space.
Open full comparisonVoyage AI's document-aware embedding model that automatically creates vectors for chunks while preserving information from the surrounding document.
Open full comparisonVoyage AI's specialist embedding model for finding relevant code from natural-language questions or other source-code context.
Open full comparisonMistral's straightforward hosted text embedding model for semantic search, clustering, classification, and RAG.
Open full comparison