AI API PricesMistral › Mistral Small 4

Mistral Small 4 pricing

Mistral Small 4 API cost from Mistral — $0.150 per 1M input tokens and $0.600 per 1M output tokens. Estimate your real monthly bill below, then compare it against every other model.

Input price
$0.150
per 1M input tokens
Output price
$0.600
per 1M output tokens
Context window
128K
tokens per request
Reference call
$0.0007
1K in + 1K out

Fast / cheap 128K context · all prices USD per 1M tokens, list price last verified 2026-06-27.

Mistral Small 4 capabilities & limits

Representative specs for this tier — confirm exact limits in the Mistral docs.

Context window128K
Max output16K
Knowledge cutoff2025
API formatOpenAI-compatible
✓ Vision✓ Tools / functions✓ JSON / structured✓ Batch API

API acting up? Check Mistral status ↗ · all AI API status

Call Mistral Small 4 from your code

The exact model string and endpoint to use. Works with the OpenAI SDK — just point base_url at the URL below.

Model stringmistral-small-latest
Base URLhttps://api.mistral.ai/v1
API key envMISTRAL_API_KEY
curl https://api.mistral.ai/v1/chat/completions \
  -H "Authorization: Bearer $MISTRAL_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "mistral-small-latest",
    "messages": [{"role": "user", "content": "Hello"}]
  }'

Model strings follow each provider's naming convention; confirm the current id in the Mistral docs before relying on it.

What a Mistral Small 4 API call actually costs

Blended input + output cost at Mistral Small 4's list rates — useful yardsticks before you commit to volume.

$0.0007
Short call
1K in · 1K out
$0.0008
Chat turn
3K in · 500 out
$0.0036
RAG / long prompt
20K in · 1K out
$0.750
1M in + 1M out
bulk job

Mistral Small 4 cost calculator

Enter your monthly volume to estimate the Mistral Small 4 bill. Runs entirely in your browser.

Estimated monthly cost

For the full multi-model breakdown, use the cost calculator.

Mistral models — Mistral Small 4 in context

How Mistral Small 4 prices against its siblings. Cheapest reference cost first.

Model Input /1M Output /1M Cost / call* Context Type
Mistral Small 4 $0.150 $0.600 $0.0007 128K Fast / cheap
Codestral
Code-specialised
$0.300 $0.900 $0.0012 32K Balanced
Mistral Large 3
Flagship
$0.500 $1.50 $0.0020 256K Flagship
Mistral Medium 3.5 $1.50 $7.50 $0.0090 256K Balanced

*Reference cost of one call with 1,000 input + 1,000 output tokens — a neutral yardstick. Use the calculator for your real usage. Cheapest row highlighted.

Cheaper than Mistral Small 4? In the same fast / cheap tier, Gemini 1.5 Flash-8B runs $0.037 in / $0.150 out. The overall cheapest model tracked is Gemini 1.5 Flash-8B — see the cheapest LLM API guide.

How Mistral Small 4 pricing works

Mistral Small 4 is billed per token: a lower input price of $0.150 per million tokens for everything you send, and a higher output price of $0.600 per million for everything the model generates. To turn that into a real Mistral Small 4 bill, multiply by your monthly request volume and average prompt/response length — the calculator above does exactly that.

Mistral Small 4 does not list a separate cached-input rate, so every input token bills at $0.150 per million. If repeated context is a big part of your prompts, a model with prompt caching may end up cheaper in practice.

List prices move fast — always confirm the current numbers on the official Mistral pricing page before relying on them for budgeting.

Mistral Small 4 pricing FAQ

How much does the Mistral Small 4 API cost?

Mistral Small 4 costs $0.150 per 1M input tokens and $0.600 per 1M output tokens. A typical 3K-input / 500-output call works out to about $0.0008.

What is Mistral Small 4's context window?

Mistral Small 4 has a 128K-token context window. Tokens you put in the prompt are billed at the $0.150 per 1M input rate.

Does Mistral Small 4 charge extra for output tokens?

Yes. Output (generated) tokens cost $0.600 per 1M, versus $0.150 per 1M for input — about 4.0× more, which is normal for LLM APIs.

Is there a cheaper alternative to Mistral Small 4?

In the same fast / cheap tier, Gemini 1.5 Flash-8B (Google) is cheaper at $0.037 in / $0.150 out. The overall cheapest model on the market is Gemini 1.5 Flash-8B.

Compare Mistral Small 4 with other models