Mistral Embed
Mistral's straightforward hosted text embedding model for semantic search, clustering, classification, and RAG.
Mistral's straightforward hosted text embedding model for semantic search, clustering, classification, and RAG.
Plain-English overview
Mistral Embed converts single or batched text inputs into 1,024-dimensional vectors through Mistral's Embeddings API. Its simple fixed-size output and published token price make it an approachable option for conventional text retrieval.
The model is narrower than newer multimodal or document-aware alternatives: applications handle parsing, chunking, query-versus-document strategy, and storage. That simplicity can be useful when the ingestion pipeline already owns those decisions.
Category comparison
These are provider-published specifications, not Cody benchmark scores. Follow the linked sources for current limits and endpoint-specific exceptions.
Pricing & comparisons
Set your usage. Your estimate updates as you type.
Assumes 600 tokens per page, processed separately. Actual token counts vary. This covers embedding only, not storage, search, or generated answers.
Mistral Embed
Mistral AI
Estimated total (USD)
For the usage above · USD · API pricing, not a subscription
Estimates exclude taxes, tools, cache storage/writes, free allowances and custom discounts. Image estimates cover output only, not prompt or reference-image charges. Quality modes differ by model. Unlisted settings are not treated as free.
API and provider access
Availability
Generally available through Mistral's hosted Embeddings API and Studio.
Mistral's standard and optional regional-inference routes have model-specific coverage; verify Embed support for an EU- or US-pinned endpoint.
Check live availabilityData and training
Mistral says API data is not used for training by default. Standard API inputs and outputs are generally retained for 30 rolling days for abuse monitoring unless approved zero-data-retention controls apply.
This is a concise reading of the cited provider material, not legal advice. A third-party gateway can have different storage, routing, training, and residency terms from the model maker's direct API.
Read the provider policyFrequently asked questions
Mistral's straightforward hosted text embedding model for semantic search, clustering, classification, and RAG. Mistral Embed converts single or batched text inputs into 1,024-dimensional vectors through Mistral's Embeddings API. Its simple fixed-size output and published token price make it an approachable option for conventional text retrieval.
Mistral Embed was released on December 11, 2023 according to the cited provider materials.
Generally available through Mistral's hosted Embeddings API and Studio. The access routes listed in this guide are Mistral AI.
$0.10 / 1M tokens. Mistral's standard hosted API list price. Batch and dedicated-deployment economics can differ from the synchronous shared endpoint.
Generally available through Mistral's hosted Embeddings API and Studio. Mistral's standard and optional regional-inference routes have model-specific coverage; verify Embed support for an EU- or US-pinned endpoint.
Mistral says API data is not used for training by default. Standard API inputs and outputs are generally retained for 30 rolling days for abuse monitoring unless approved zero-data-retention controls apply. The policy belongs to the provider route and account terms, so verify it again before production use.
USD per million tokens for the shown API routes at standard context length. Cache and long-context rates may differ. Models without matching reviewed prices are omitted; lower cost does not mean better quality.
Mistral AI
This model
Voyage AI
Voyage AI
Voyage AI
OpenAI
Related comparisons
Google's multimodal embedding model for placing text, images, video, audio, and PDFs in one searchable vector space.
Open full comparisonOpenAI's highest-capability text embedding model for semantic search, recommendations, clustering, and retrieval pipelines.
Open full comparisonCohere's enterprise embedding model for multilingual text, images, and visually rich documents with a 128K context window.
Open full comparisonVoyage AI's quality-first general embedding model for text and code retrieval, with adjustable dimensions and a shared family vector space.
Open full comparisonVoyage AI's document-aware embedding model that automatically creates vectors for chunks while preserving information from the surrounding document.
Open full comparisonVoyage AI's specialist embedding model for finding relevant code from natural-language questions or other source-code context.
Open full comparison