Today for AI

Simon Willison 博客 · 2026/10/6 20:18:19

Mistral 发布 Large 4:万亿参数 MoE 模型预览版上线

原标题:Introducing Mistral Large 4: Le chonk
86AI 研判分
核心综述

Mistral AI 正式推出旗舰模型 Mistral Large 4 的预览版,该模型采用总参数 1 万亿、激活参数 490 亿的混合专家(MoE)架构。目前通过 API 提供访问,承诺月底开源权重,在 Artificial Analysis 基准测试中得分 38,性能略低于 DeepSeek 4.1 Flash 但较前代有显著提升。

报道全文原始报道全文

Introducing Mistral Large 4: Le chonk

Mistral are back in the game. Today they're releasing a preview of Mistral Large 4, a 1 trillion parameter, 49 billion active parameter model trained on their own cluster of 3,800 NVIDIA Grace Blackwell GPUs.

The preview is available via their API. They promise to release the open weights model at the "end of this month".

The model only supports two reasoning levels - "none" and "high" - via the Mistral API. Here are both pelicans - the "high" one looks better, though surprisingly it only used 2,717 output tokens compared to "none" which used 3,275:

It's good. The pouch is great, the bicycle frame is the right size, it has feet on pedals. Both pedals appear in front of the frame though. Nice gradients.

On Artificial Analysis it scores 38, just behind DeepSeek 4.1 Flash, which is a 552B model. It's a huge improvement on last December's Mistral Large 3, which drew this terrible pelican and scored 9 on AA.

It's certainly not a Fable-class model, but it's great to see Mistral put out a model that's back to being maybe about 6 months behind the frontier.

Via Hacker News

Tags: ai, generative-ai, llms, mistral, pelican-riding-a-bicycle, llm-release