vs Large 3

Mistral Large 4 vs Mistral Large 3

Mistral Large 4 不只是小改版,而是一大步:總參數從 6,750 億成長到 1.05 兆,加入 Large 3 沒有的可切換推理模式,API 提供的上下文加倍到 524,288 個 token,訓練資料涵蓋 160 多種語言;原價大約是 Large 3 的 2.7 倍(目前五折優惠期間約 1.4 倍)。

更新於 2026-10-06 · 作者:suifeng

Mistral Large 4 重點規格

開發商
Mistral AI(法國巴黎)
發表日期
2026 年 10 月 6 日(公開預覽版)
參數量
總計 1.05 兆,每個 token 啟用 490 億(混合專家模型 MoE)
上下文長度
API 提供 524,288 個 token
最大輸出
262,144 個 token
輸入/輸出
可輸入文字和圖片,輸出文字
API 價格(優惠中)
每 100 萬 token 輸入 $0.68/輸出 $2.09(原價 $1.36/$4.18)
API 模型名稱
mistral-large-4
推理模式
reasoning_effort: "high" 或 "none"
開放權重
預計 2026 年 10 月底釋出

資料來源:Mistral AI 官方公告與定價頁、OpenRouter 模型頁、Hugging Face 模型頁,2026 年 10 月 6 日查核。

從 Large 3 到 Large 4 改變了什麼

兩次發表相隔十個月:Mistral Large 3(2512)在 2025 年 12 月 2 日推出,Mistral Large 4 在 2026 年 10 月 6 日進入公開預覽。最主要的改變是規模、推理和上下文。

  • 規模:總參數 1.05 兆、每個 token 啟用 490 億,上一代是總參數 6,750 億、啟用 410 億(總參數 +56%,啟用參數 +20%)。
  • 推理:Large 4 是指令與推理合一的混合模型,用 reasoning_effort("high" 或 "none")控制;Large 3 的模型卡則說明它不是專門的推理模型。
  • 上下文:API 提供 524,288 個 token(Mistral 文件寫 1M),上一代是 256K。
  • 語言:訓練資料涵蓋超過 160 種語言,包含所有歐盟官方語言;Large 3 的模型卡列出的是數十種語言。
  • 視覺:兩者都能讀圖;Large 4 使用 16 億參數的視覺編碼器,Large 3 則是 25 億參數。
  • 授權:Large 3 採用 Apache 2.0;Mistral 還沒公布 Large 4 的授權條款。

規格比較表

除了視覺編碼器的大小,Large 4 在每一項容量數字上都勝出;而 Large 3 目前的優勢是權重已經可以下載。

規格Mistral Large 4Mistral Large 3(2512)
發表日期2026 年 10 月 6 日(公開預覽)2025 年 12 月 2 日
總參數/啟用參數1.05T / 49B(含嵌入層 52B)675B / 41B
視覺編碼器1.6B 參數2.5B 參數
上下文長度API 提供 524,288 token(文件寫 1M)256K(262,144 token)
最大輸出262,144 token209,715 token(OpenRouter)
推理模式有,reasoning_effort 設 high 或 none非專門的推理模型
語言訓練資料涵蓋 160+ 種數十種
權重預計 2026 年 10 月底釋出已提供:FP8、NVFP4、BF16
授權條款尚未公布Apache 2.0
資料來源:Mistral Docs、Mistral Large 4 發表文章、Hugging Face 模型卡、OpenRouter,2026 年 10 月 6 日查核。

每 100 萬 token 的價差

以原價計算,Mistral Large 4 大約是 Mistral Large 3 的 2.7 到 2.8 倍;在 Mistral 限時五折期間,只有約 1.4 倍。倍數的算法是 Large 4 價格除以 Large 3 價格。

每 100 萬 token(美元)Large 4 原價Large 4 優惠價Large 3原價倍數優惠價倍數
輸入$1.36$0.68$0.502.72×1.36×
快取輸入$0.14$0.07$0.052.80×1.40×
輸出$4.18$2.09$1.502.79×1.39×
批次輸入/輸出$0.68 / $2.09$0.34 / $1.045$0.25 / $0.752.72× / 2.79×1.36× / 1.39×
資料來源:Mistral Docs 定價頁;倍數為本站計算,2026 年 10 月 6 日查核。

一般每月工作量在兩個模型各要多少錢

每月輸入 3,000 萬、輸出 1,000 萬 token 的工作量,用 Large 3 是 $30.00,用 Large 4 優惠價是 $41.30,Large 4 原價則是 $82.60。

計算方式:Large 3 = 30 × $0.50 + 10 × $1.50 = $30.00;Large 4 優惠價 = 30 × $0.68 + 10 × $2.09 = $41.30;Large 4 原價 = 30 × $1.36 + 10 × $4.18 = $82.60。推理產生的內容也算輸出 token,所以簡單的提示如果還開著 reasoning_effort "high",Large 4 的帳單會更高;Artificial Analysis 評 Large 4「非常囉嗦」(跑完指數用了 200M 輸出 token,中位數是 81M)。

如果你只是要用來聊天,完全不用算 token:我們的 Pro 方案每月 US$39.90 或每年 US$199.90,每月 3,000 點,1 點大約是一則簡短回覆。

對這個主題還有疑問?

直接問 Mistral Large 4。免費帳號每天 15 點。

馬上問

品質提升:公開數字看得出什麼

Mistral 的 Large 4 發表圖表沒有放 Large 3,而 Large 3 也不在目前 44 個模型的 Vals Index 裡,所以兩者沒有共同的跑分項目可比。Mistral 公布的是 Large 4 和自家另一個現行模型 Mistral Medium 3.5 的比較,差距相當大。

以原價來看,Large 4 甚至比 Medium 3.5 便宜:每 100 萬 token $1.36/$4.18,Medium 3.5 是 $1.50/$7.50。

基準測試Mistral Large 4Mistral Medium 3.5
AutomationBench(657 個工作流程)59.9%6.3%
Finance Agent v254.732.1
Finch(FinWorkBench)67.436.7
SciCode-Verified pass@191.870
ChartQA Pro63.155.4
資料來源:Mistral Large 4 發表文章(廠商自報),2026 年 10 月 6 日查核。

自行架設:Large 3 現在就能跑,Large 4 需要更多記憶體

Large 3 的 FP8 權重約 675 GB,放在一台 8× H200 節點(1,128 GB)上還剩約 453 GB;Large 4 的 FP8 權重約 1,050 GB,在同一台節點上只剩約 78 GB。4-bit 的話,Large 3 約需 338 GB、Large 4 約需 525 GB,兩者都放得進 8× H100(640 GB)。

Large 3 現在就能在 vLLM 1.12.0 以上版本執行。目前還沒有推論引擎宣布支援 Large 4;完整的記憶體計算請看「本地部署」指南。

把 API 呼叫從 Large 3 換成 Large 4

換成 Large 4,大致上只要在同一個 chat completions 端點把模型名稱改成 mistral-large-4(別名 mistral-large-4-0)。有兩個地方需要改程式:reasoning_effort 設為 "high" 時,訊息內容會從字串變成思考區塊和文字區塊組成的清單;另外,Mistral 文件要求在多輪對話的歷史紀錄中,送回完整的助理訊息,包含思考區塊。

把 reasoning_effort 設成 "none",就能維持一般字串格式的回應,原本解析 Large 3 回應的程式碼可以直接沿用。

# Set API_KEY to your Mistral API key first
curl https://api.mistral.ai/v1/chat/completions \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer $API_KEY" \
  -d '{"model":"mistral-large-4","messages":[{"role":"user","content":"Summarize this contract clause."}],"reasoning_effort":"none"}'

資料來源

FAQ

vs Large 3 常見問題

延伸閱讀

更多 Mistral Large 4 資訊

拿一個真正的問題問 Mistral Large 4

貼上一段出錯的程式、一條合約條款,或一段亂糟糟的試算表備註。免費帳號每天送 15 點。

開始聊天