Mistral 发布 Large 4 预览版
Mistral 推出万亿参数模型 Large 4,性能大幅提升,月底将开源权重。
Mistral 发布 Large 4 预览版,拥有 1 万亿参数,活跃参数 490 亿。该模型在 3,800 台 NVIDIA Grace Blackwell GPU 上训练,通过 API 提供。在 Artificial Analysis 基准上得分为 38,落后于 DeepSeek 4.1 Flash 的 552B 模型。相比去年 12 月的 Large 3 版本(得分 9)有显著提升。
Introducing Mistral Large 4: Le chonk Mistral are back in the game. Today they're releasing a preview of Mistral Large 4, a 1 trillion parameter, 49 billion active parameter model trained on their own cluster of 3,800 NVIDIA Grace Blackwell GPUs. The preview is available via their API. They promise to release the open weights model at the "end of this month". The model only supports two reasoning levels - "none" and "high" - via the Mistral API. Here are both pelicans - the "high" one looks better, though surprisingly it only used 2,717 output tokens compared to "none" which used 3,275: On Artificial Analysis it scores 38 , just behind DeepSeek 4.1 Flash, which is a 552B model. It's a huge improvement on last December's Mistral Large 3, which drew this terrible pelican and scored 9 on AA . It's certainly not a Fable-class model, but it's great to see Mistral put out a model that's back to being maybe about 6 months behind the frontier. Via Hacker News Tags: ai , generative-ai , llms , mistral , pelican-riding-a-bicycle , llm-release