About Mistral Large 3
Mistral Large 3 is the flagship of Mistral's open-weight range: an Apache 2.0-licensed, granular mixture-of-experts model with a vision encoder built in, which its model card sums up as a state-of-the-art general-purpose multimodal system. The active-parameter slice per token is a small fraction of the very large total, and latent-compression attention keeps its long context affordable in memory.
Mistral pitches it at long-document understanding, daily-driver assistants, agentic work and coding — flagship duties — with the unusual property that the same weights are downloadable, so nothing about it is locked to Mistral's own API.
Hosted, Large 3 is priced like a mid-tier model despite flagship positioning — Mistral's list rates undercut the American frontier tiers by a wide margin, and the derived cache discount applies as across the platform. The open weights create a real build-versus-buy decision at volume: steady high-throughput workloads can amortise self-hosting on rented accelerators, while spiky usage stays cheaper on the API. In the rivals table above it tends to sit near much smaller closed models on price, which is the argument in a sentence: capability from the open flagship at rates set against the mid-market.
Frequently asked questions
How much does the Mistral Large 3 API cost?
Mistral Large 3 costs $0.5 per million input tokens and $1.5 per million output tokens, with cached input at $0.05 per million (90% off). Note: Cached input derived from Mistral’s stated −90% discount.
How much does a typical chat message cost with Mistral Large 3?
A short chat message (about 500 input and 250 output tokens) costs roughly $0.0006 with Mistral Large 3. At 1,000 such requests a day, that is about $18.75 per month.
Is Mistral Large 3 cheap compared to similar models?
Mistral Large 3 has the #9 cheapest input rate of the 29 models we track. On a mixed workload (10K input / 2K output), it costs $0.0080 per request, versus $0.0065 for GPT-5 mini.
What does Mistral Large 3 cost per 1,000 tokens?
Per 1,000 tokens, Mistral Large 3 costs $0.00050 for input and $0.0015 for output, or $0.00005 for cached input. Providers quote rates per million tokens ($0.5/M in, $1.5/M out here), so divide by a thousand for the per-1K figure older pricing pages used.
How large a prompt can Mistral Large 3 take?
Mistral Large 3 accepts up to 256K tokens of input context. At its input rate, filling the entire 256K-token window costs about $0.131 per request before any output.
Related
Compare all models at once ·Count tokens in your prompt · Run Mistral Large 3 locally — VRAM requirements
All Mistral model costs: Mistral Medium 3.5 · Mistral Small 4
Last updated 2026-09-03. Prices verified against Mistral's official pricing page; see the methodology.