Mistral Large 4 API Pricing (Verified Oct 6, 2026)

The Mistral Large 4 API lists at $1.36 per 1M input tokens, $4.18 per 1M output tokens and $0.14 per 1M cached input tokens, and Mistral AI is currently selling it at a temporary 50% discount: $0.68 input, $2.09 output and $0.07 cached. Use the calculator below to turn those rates into a monthly bill for your own traffic.

Updated 2026-10-06

Mistral Large 4 costs $92.80 per month ($0.0019 per request)

ModelInput $/MOutput $/MPer month
Mistral Large 3$0.50$1.50$67.50
DeepSeek V4 Pro (off-peak)$0.66$1.98$89.10
Mistral Large 4$0.68$2.09$92.80
DeepSeek V4 Pro (peak)$1.32$3.96$178
Mistral Large 4 (list price)$1.36$4.18$186
GLM-5.3$1.40$4.40$193
Qwen3.8 Max$2.00$6.00$270
Claude Sonnet 5.5$2.00$10.00$350
GPT-6.1 Sol$2.00$10.00$350
Claude Opus 5.5$4.00$20.00$700
GPT-6 Astra$10.00$50.00$1,750

Prices per 1M tokens from each provider’s pricing page and OpenRouter, checked Oct 6, 2026. Mistral Large 4 is on a 50% launch sale.

Mistral Large 4 at a glance

Developer
Mistral AI (Paris, France)
Released
October 6, 2026 (public preview)
Parameters
1.05T total, 49B active per token (Mixture of Experts)
Context window
524,288 tokens via the API
Max output
262,144 tokens
Input / output
Text and images in, text out
API price (sale)
$0.68 input / $2.09 output per 1M tokens (list $1.36 / $4.18)
API model name
mistral-large-4
Reasoning
reasoning_effort: "high" or "none"
Open weights
Scheduled for the end of October 2026

Sources: Mistral AI announcement and pricing page, OpenRouter model listing, Hugging Face model page. Checked Oct 6, 2026.

Mistral Large 4 price per million tokens

Right now you pay half the list price. Mistral’s pricing page shows $1.36 / $4.18 struck through and labels $0.68 / $2.09 as a temporary sale price. The discount is exactly 50% on every line: 0.68 ÷ 1.36, 2.09 ÷ 4.18 and 0.07 ÷ 0.14 all equal 0.50.

Cached input is the cheapest line on the bill. A cached token costs about one tenth of a fresh input token ($0.07 versus $0.68 at the sale price), so long system prompts and repeated documents get much cheaper on the second call.

Token typeList price (per 1M)Current sale price (per 1M)
Input$1.36$0.68
Cached input$0.14$0.07
Output$4.18$2.09
Source: Mistral Docs pricing page and Mistral Large 4 model card, checked Oct 6, 2026.

Batch, priority and regional tiers

Batch is the cheapest way to run Mistral Large 4: it takes another 50% off, so at the current sale price a million input tokens costs $0.34 and a million output tokens costs $1.045. Priority processing adds 75% for faster, reserved capacity, and regional inference (for example the Mistral-operated European deployment) adds 10%.

All four tiers carry the same temporary 50% sale, so the gap between them stays the same whichever price applies when you read this.

TierInput list / saleCached list / saleOutput list / sale
Standard$1.36 / $0.68$0.14 / $0.07$4.18 / $2.09
Batch (−50%)$0.68 / $0.34$0.07 / $0.035$2.09 / $1.045
Priority (+75%)$2.38 / $1.19$0.245 / $0.1225$7.315 / $3.6575
Regional inference (+10%)$1.496 / $0.748$0.154 / $0.077$4.598 / $2.299
Source: Mistral Docs pricing page (per 1M tokens, USD), checked Oct 6, 2026.

Why OpenRouter and Vercel show $0.68 / $2.09

OpenRouter and Vercel AI Gateway pass through Mistral’s own sale price, so the per-token rate is the same on all three. OpenRouter’s endpoint data for mistralai/mistral-large-4-0 carries a discount value of 0.5, and Mistral is the only provider serving the model there.

Mistral has not announced when the sale ends. Some summaries claim the discount runs for the first two weeks, but no Mistral page states a date. Benchmark sites such as Artificial Analysis and Vals AI calculate their cost figures at the $1.36 / $4.18 list price, so budget at list price if you are planning beyond this month.

Where you buyModel IDInput / output per 1MCached input per 1M
Mistral Studio APImistral-large-4$0.68 / $2.09 (list $1.36 / $4.18)$0.07
OpenRoutermistralai/mistral-large-4-0$0.68 / $2.09$0.07
Vercel AI Gatewaymistral/mistral-large-4$0.68 / $2.09not listed
Source: Mistral Docs, OpenRouter endpoints API, Vercel AI Gateway model page, checked Oct 6, 2026.

What real workloads cost

Output tokens drive most bills, because they cost about three times as much as input. The examples below use the formula tokens ÷ 1,000,000 × price and assume 30 days a month; plug your own numbers into the calculator above for a closer estimate.

Reasoning makes output longer. With reasoning_effort set to "high", the model returns thinking chunks before its answer, and Artificial Analysis recorded 200M output tokens for its full test run against a median of 81M across models, so it rates Mistral Large 4 as very verbose. Setting reasoning_effort to "none" for simple tasks keeps output short.

WorkloadAt sale priceAt list price
One call: 10K input + 2K output$0.011$0.022
One long-document summary: 100K input + 2K output$0.072$0.144
Filling the full 524,288-token context once (input only)$0.36$0.71
A maximum 262,144-token answer (output only)$0.55$1.10
Chatbot: 1,000 requests/day, 2K in + 500 out, no cache$72.15 / month$144.30 / month
Same chatbot with 70% of input served from cache$46.53 / month$93.06 / month
Coding agent: 200 calls/day, 50K in + 5K out$266.70 / month$533.40 / month
10M input + 2M output via Batch$5.49$10.98
Source: computed from Mistral Docs pricing (per 1M tokens), checked Oct 6, 2026.

Mistral Large 4 price vs DeepSeek V4, GPT-6, Claude and others

At the sale price, Mistral Large 4 costs about the same as DeepSeek V4 Pro at off-peak hours and far less than the US flagship models. At list price it is 2.9× cheaper on input and 4.8× cheaper on output than Claude Opus 5.5, and 7.4× / 12× cheaper than GPT-6 Astra. Inside Mistral’s own lineup it costs more than Mistral Large 3 ($0.50 / $1.50) but less per output token than Mistral Medium 3.5 ($7.50).

The last column applies one fixed workload (10M input + 2M output tokens) to every model, which makes the gap easy to read. For score-for-score comparisons, see the benchmarks page and the /vs-deepseek-v4 and /vs-mistral-large-3 guides.

ModelInput per 1MOutput per 1MCost of 10M in + 2M out
Mistral Large 4 (sale)$0.68$2.09$10.98
Mistral Large 4 (list)$1.36$4.18$21.96
Mistral Large 3$0.50$1.50$8.00
Mistral Medium 3.5$1.50$7.50$30.00
Mistral Small 4$0.15$0.60$2.70
DeepSeek V4 Pro 0813 (off-peak)$0.66$1.98$10.56
DeepSeek V4 Pro 0813 (peak)$1.32$3.96$21.12
GLM-5.3 (Z.AI)$1.40$4.40$22.80
Qwen3.8 Max$2.00$6.00$32.00
GPT-6.1 Sol (short context)$2.00$10.00$40.00
Claude Sonnet 5.5$2.00$10.00$40.00
Gemini 3.1 Pro Preview (≤200K)$2.00$12.00$44.00
Kimi K3$3.00$15.00$60.00
Claude Opus 5.5$4.00$20.00$80.00
GPT-6 Astra (short context)$10.00$50.00$200.00
Source: Mistral, DeepSeek, OpenAI, Anthropic and Google pricing pages plus OpenRouter model API; last column computed, checked Oct 6, 2026.

Cost per benchmark task

Independent testers put a price on whole tasks, not just tokens. Artificial Analysis reports $1.13 per Intelligence Index task for Mistral Large 4 at list price, and Vals AI reports $13.78 per test on its Vals Index. Both include the extra output from reasoning, so they are a realistic upper bound for hard, multi-step work.

EvaluatorMetricMistral Large 4 result
Artificial AnalysisCost per Intelligence Index task (list price)$1.13
Artificial AnalysisOutput tokens used for the full index200M (median 81M)
Vals AICost per test, Vals Index v2.1$13.78
Source: Artificial Analysis and Vals AI model pages, checked Oct 6, 2026.

Five ways to cut your Mistral Large 4 bill

The biggest savings come from batch jobs and caching; the rest come from controlling output length.

  • Send offline jobs (evaluations, bulk classification, document backfills) to the Batch endpoint /v1/batch for 50% off every token type.
  • Keep long system prompts and shared documents identical at the start of each request so they can be billed at the cached rate of $0.07 instead of $0.68 per 1M.
  • Use reasoning_effort "none" for lookups, rewrites and extraction; keep "high" for code, math and multi-step agents where the thinking pays off.
  • Cap max_tokens on endpoints that only need short answers, because each 1M output tokens costs $2.09 at the sale price.
  • Route simple traffic to Mistral Small 4 ($0.15 / $0.60) and keep Mistral Large 4 for the requests that need its coding, legal and vision strength.

Prefer a flat monthly price?

If you want to use Mistral Large 4 without an API key or token math, chat with it on our homepage. Anyone gets 3 free messages a day without an account, a free account gets 15 credits a day, and Pro costs $39.90 a month or $199.90 a year for 3,000 credits a month, where one credit is about one short reply. Plan details are on the /pricing page, and the /free guide lists every free way to try the model.

Frequently asked questions

More about Mistral Large 4