Mistral Large 4 在 Code Arena WebDev 榜单排名第 45,得分 1534,比 Mistral Large 3 的第 130 名提升 304 分,也比 Mistral Medium 3.5 的第 122 名高 271 分。
Mistral Large 4 在 Code Arena WebDev 的排名与价格对比,可帮助读者判断其相对前沿闭源模型的位置。
Mistral Large 4 by @MistralAI just landed in the Code Arena: WebDev at #45. With 1534 pts, this is a +304 pt improvement from Mistral Large 3 at #130!
This release also marks a 271-point improvement over its next-best-performing variant, Mistral Medium 3.5, at #122.
Mistral Large 4 delivers performance within just 2 points of Claude Opus 4.8 (High) at a nearly 6× lower blended price: $3.47/M vs. $20/M tokens. Nearby but higher-performing models are available for less, keeping it just off the Code Arena: WebDev Pareto frontier.
@MistralAI announced its open weights will be released at the end of October. Stay tuned for its score on Arena within open models, and congrats to the team on this release!
Meet Mistral Large 4, aka Le Chonk. • 1T parameters, natively multimodal. 49B active. It is the best open weights model from US or Europe on aggregated benchmarks. • State-of-the-art on critical workloads, including cyber defense, manufacturing and finance and it surpasses closed frontier models on visual grounding. • Forged in Europe end-to-end and is deployable from Europe via our own Mistral Cloud infrastructure. • Available to all via API today. Working with cybersecurity partners privately. Open weights release end of October.在 X 查看被引用的帖子
来源:Arena.ai · x.com