█

AI/TLDR

Mistral Large 4

Mistral's 1T-parameter multimodal MoE flagship with 49B active parameters and a 1M-token context

Mistral LargeAPI onlyPublic preview
Released
6 Oct 2026
Context
1M
Parameters
1.05T total · 49B active
Input
$1.36 / 1M tokens
License
Not yet published (open weights announced for end of October 2026)
Coverage
1 story

Overview

Mistral Large 4 is Mistral AI's flagship model, launched as a public preview on October 6, 2026. It is a natively multimodal mixture-of-experts (MoE) model with 1.05 trillion total parameters, 49 billion active parameters and a 1.6B vision encoder, and Mistral's documentation lists a 1M-token context window. Mistral calls it "ML4" unofficially, and "le Chonk" very officially.

Mistral trained the model on 3,800 NVIDIA Grace Blackwell GPUs in its own European datacenters. The launch post focuses on coding, agentic workflows and multimodal understanding, with extra attention on cybersecurity, finance and legal work, and says the model supports more than 160 languages, including all official EU languages. It reports 61.7% on DeepSWE v1.1, 59.4% on SWE-Atlas-QnA, 28.3% on Terminal-Bench 4, 59.9% on AutomationBench and 93% of the 40 Cybench challenges.

The preview API is available on Mistral Studio under the model id mistral-large-4, with structured outputs, function calling, document Q&A and batching. Mistral says the weights drop at the end of October 2026; the license for those weights is not yet stated.

Released2026-10-06
LicenseNot yet published (open weights announced for end of October 2026)
WeightsAPI only
Parameters1.05T total · 49B active
Context1M
ArchitectureMixture-of-Experts
ModalitiesText, Vision
StatusPublic preview

Benchmarks

Surge AI human evaluation of coding quality (score out of 5), as published by Mistral

BenchmarkMistral Large 4 PreviewClaude Opus 5GLM-5.3Kimi K3GLM-5.2
Surge AI coding human eval3.74 /54.22 /53.6 /53.59 /53.4 /5

Comparison source ↗

Pricing

Input$1.36 / 1M tokens
Cached input$0.14 / 1M tokens
Output$4.18 / 1M tokens

List price; the docs page shows a discounted $0.68 input / $0.07 cached / $2.09 output during the preview

Pricing source ↗

Strengths

  • Frontier-scale MoE: 1.05T total parameters with only 49B active per token
  • 1M-token context window on the preview API
  • Native vision through a 1.6B vision encoder
  • Strong security results: 93% of Cybench challenges and 82% on the AA Cyber Index vulnerability-reproduction test
  • Open weights announced for the end of October 2026

Best for

  • Coding agents and long terminal workflows
  • Agentic business automation
  • Security work such as vulnerability reproduction
  • Finance and legal document analysis across many languages

How to access

ProviderModel ID
Mistral AI (Mistral Studio) ↗mistral-large-4

Mistral Large — every version

The full lineage of the Mistral Large line, newest first. Every version has its own page — click any to compare specs, benchmarks and pricing.

VersionReleasedContextLicense
Mistral Large 42026-10-061M—
Mistral Large 3current2025-12-02256KApache-2.0
Mistral Large 2.1 (24.11)2024-11-18—Open weights
Mistral Large 2 (24.07)2024-07-24—Open weights
Mistral Large (24.02)2024-02-26—Proprietary

FAQ

Is Mistral Large 4 open-weight?

Mistral Large 4 is described in Mistral's docs as an open-weight model, but at launch on October 6, 2026 only the preview API was live. Mistral's launch post says the weights drop at the end of October 2026. The license for the weights has not been published yet.

How much does Mistral Large 4 cost?

Mistral Large 4 lists at $1.36 per million input tokens, $0.14 per million cached input tokens and $4.18 per million output tokens. During the public preview, Mistral's docs page shows discounted rates of $0.68 input, $0.07 cached input and $2.09 output per million tokens.

How big is Mistral Large 4?

Mistral Large 4 is a mixture-of-experts model with 1.05 trillion total parameters and 49 billion active parameters per token, plus a 1.6B vision encoder. It is larger than Mistral Large 3 (675B total, 41B active) and its context window grows from 256K to 1M tokens.