AI Trend Notifier
EN한
← wiki

$ cat wiki/models/mistral-large-4.md

Mistral Large 4

modelupdated 2026-10-07created 2026-10-07

Compared with

Mistral's multimodal API preview, announced on 2026-10-06. The announcement promises weights by the end of October; it does not establish that weights are available now. (source)

Spec

AttributeValue
DeveloperMistral AI
Released2026-10-06 (API public preview)
Announced2026-10-06
Context window1M
Pricing$1.36/M input · $4.18/M output (announcement)
Licenseunknown
AvailabilityMistral Studio API preview; weights promised by end of October 2026
Release, pricing and availability follow the announcement; context follows the model documentation, which identifies mistral-large-4. The sources read do not name a weight licence. (announcement) (documentation)

Release Date

The API preview launched on 2026-10-06. Mistral says red-teaming with vetted partners and state authorities precedes the weight release. Future open weights are a commitment, not a completed release. (source)

Benchmarks

Mistral reports DeepSWE v1.1 61.7%, SWE-Atlas-QnA 59.4%, Terminal-Bench 4 28.3%, and AutomationBench 59.9% across 657 business workflows. These are figures from the vendor's announcement, not an independently reproduced evaluation. (source)

Its cyber comparisons include models that refuse tasks. Interpretation: access policy can affect the ordering, so these figures should not be read as capability measured under identical safeguards. (source)

Use Cases

The announcement targets coding agents, multimodal document work and enterprise security. It describes shared RL environments and verification components spanning chat, scientific tasks and long-horizon tool use. (source)

Compared To

Mistral's blind coding evaluation with Surge AI assigns ML4 Preview 3.74 on a 1–5 scale, versus 4.22 for Claude Opus 5 and 3.60 for GLM-5.3. This is a separate human evaluation, not the coding benchmark aggregate. (source)

Conflicting Reports

The announcement describes 1 trillion total and 49 billion active parameters; the model documentation says 1.05T total, 52B active and a 1.6B vision encoder. Neither source read reconciles these descriptions. Both are retained without choosing a parameter count. (announcement) (documentation)

The documentation displays both $1.36/$4.18 and $0.68/$2.09 input/output pairs, plus $0.14/$0.07 cached-input values, without an unambiguous rate-mode label in the captured text. The Spec table therefore attributes its prices to the announcement; the lower pair is not adopted as a price cut. (documentation)

Referenced by

Sources