Mistral Large 4 (Le Chonk) pricing
Mistral Large 4 (Le Chonk) API cost from Mistral — $0.680 per 1M input tokens and $2.09 per 1M output tokens. Public preview · Standard sale rates · promotion end date not published. Estimate your real monthly bill below, then compare it against every other model.
Flagship 1M context · all prices USD per 1M tokens, this price verified 2026-10-07.
Mistral Large 4 (Le Chonk) capabilities & limits
Representative specs for this tier — confirm exact limits in the Mistral docs.
API acting up? Check Mistral status ↗ · all AI API status
Call Mistral Large 4 (Le Chonk) from your code
The exact model string and endpoint to use. Works with the OpenAI SDK — just point base_url at the URL below.
mistral-large-4https://api.mistral.ai/v1MISTRAL_API_KEYcurl https://api.mistral.ai/v1/chat/completions \
-H "Authorization: Bearer $MISTRAL_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "mistral-large-4",
"messages": [{"role": "user", "content": "Hello"}]
}' Model strings follow each provider's naming convention; confirm the current id in the Mistral docs before relying on it.
What a Mistral Large 4 (Le Chonk) API call actually costs
Blended input + output cost at Mistral Large 4 (Le Chonk)'s current Standard sale rates — useful yardsticks before you commit to volume.
Mistral Large 4 (Le Chonk) cost calculator
Enter your monthly volume to estimate the Mistral Large 4 (Le Chonk) bill. Runs entirely in your browser.
Peak/off-peak and long-context pricing are applied automatically where published. Cache-write, Batch/Flex and Fast-mode adjustments are not included. For the full multi-model breakdown, use the cost calculator.
Mistral models — Mistral Large 4 (Le Chonk) in context
How Mistral Large 4 (Le Chonk) prices against its siblings. Cheapest reference cost first.
| Model | Input /1M | Output /1M | Cost / call* | Context | Type |
|---|---|---|---|---|---|
| Mistral Small 4 | $0.150 | $0.600 | $0.0007 | 128K | Fast / cheap |
| Codestral Code-specialised | $0.300 | $0.900 | $0.0012 | 32K | Balanced |
| Mistral Large 3 Flagship | $0.500 | $1.50 | $0.0020 | 256K | Flagship |
| Mistral Large 4 (Le Chonk) Public preview · Standard sale rates · promotion end date not published | $0.680 | $2.09 | $0.0028 | 1M | Flagship |
| Mistral Medium 3.5 | $1.50 | $7.50 | $0.0090 | 256K | Balanced |
*Reference cost of one call with 1,000 input + 1,000 output tokens — a neutral yardstick. Use the calculator for your real usage. Cheapest row highlighted.
What is Le Chonk AI?
Le Chonk is Mistral Large 4 (ML4), the general-purpose multimodal model that Mistral introduced as a public preview on October 6, 2026. The nickname and model name refer to the same entry here. It supports document questions, structured outputs and function calling, with a documented 1M-token context window.
The release announcement offers preview API access through Mistral Studio and says weights are planned for the end of October. As of this verification, that announcement is a release plan, not confirmation that downloadable weights are already available. We have not independently benchmarked the model.
Mistral Large 4 sale price vs original price
| USD per 1M tokens | Current Standard sale | Official original price |
|---|---|---|
| Input | $0.680 | $1.36 |
| Cached input | $0.070 | $0.14 |
| Output | $2.09 | $4.18 |
The official pricing table labels these as sale and original prices. Do not assume the sale is permanent or that the original rates will resume on a particular date. A request using 10,000 uncached input tokens and 2,000 output tokens costs $0.01098 at the sale rates: (10,000 × $0.68 + 2,000 × $2.09) ÷ 1,000,000. At the displayed original rates, the same token usage would cost $0.02196.
These examples cover model tokens only. Additional requests, retries, external tools and infrastructure can increase the total cost. A lower API bill does not establish equivalent task quality; compare outputs on your own workload.
Le Chonk API access and limits
Use mistral-large-4, the identifier displayed on the official model page, with the Mistral API. “Le Chonk AI” is a search alias, not the API identifier. The copyable request above uses the documented model name.
The model remains in public preview. The source page does not publish a separate maximum output-token limit or knowledge cutoff; this page leaves those fields unconfirmed instead of inferring them from other Mistral models. The API price is separate from any Le Chat subscription.
How Mistral Large 4 (Le Chonk) pricing works
Mistral Large 4 (Le Chonk) is billed per token: a lower input price of $0.680 per million tokens for everything you send, and a higher output price of $2.09 per million for everything the model generates. To turn that into a real Mistral Large 4 (Le Chonk) bill, multiply by your monthly request volume and average prompt/response length — the calculator above does exactly that.
Prompt caching is the biggest lever on your Mistral Large 4 (Le Chonk) cost: cached input tokens bill at just $0.070 per million — about 90% cheaper. If you reuse a long system prompt or context across calls, caching can dramatically cut the input portion of your bill.
List prices move fast — always confirm the current numbers on the official Mistral pricing page before relying on them for budgeting.
Mistral Large 4 (Le Chonk) pricing FAQ
How much does the Mistral Large 4 (Le Chonk) API cost?
Mistral Large 4 (Le Chonk) costs $0.680 per 1M input tokens and $2.09 per 1M output tokens, with cached input at $0.070 per 1M (about 90% off). A typical 3K-input / 500-output call at the displayed rate works out to about $0.0031.
What is Mistral Large 4 (Le Chonk)'s context window?
Mistral Large 4 (Le Chonk) has a 1M-token context window. Standard uncached input is $0.680 per 1M tokens.
Does Mistral Large 4 (Le Chonk) support prompt caching?
Yes. At Standard rates, Mistral Large 4 (Le Chonk) bills cache-read tokens at $0.070 per 1M instead of $0.680 uncached — roughly 90% cheaper for the repeated part of your prompts.
Is there a cheaper alternative to Mistral Large 4 (Le Chonk)?
In the same flagship tier, MiniMax-M3 (MiniMax) is cheaper at $0.300 in / $1.20 out. For the 1K-input / 1K-output reference workload, the lowest-cost model tracked in this catalog is Gemini 1.5 Flash-8B.