Peaked at #1 Off the board
Mistral releases Large 4 open model
Mistral launches Large 4, a 1T-parameter natively multimodal model with 49B active, calling it the best open-weights model from US or Europe.
Key points
- Mistral AI released Mistral Large 4, aka Le Chonk, with 1T parameters, 49B active, and natively multimodal.
- Mistral AI said it is the best open weights model from US or Europe on aggregated benchmarks, with open weights release end of October.
- Mistral AI said the model is forged in Europe end-to-end and deployable from Europe via its own Mistral Cloud infrastructure, and available to all via API today.
- Arthur Mensch said "We're building AGI from this train station."
Key points and the reaction summary are written by AI from the posts on this page. Check the original post. How we use AI
Original post
Meet Mistral Large 4, aka Le Chonk.
• 1T parameters, natively multimodal. 49B active.
It is the best open weights model from US or Europe on aggregated benchmarks.
• State-of-the-art on critical workloads, including cyber defense, manufacturing and finance and it surpasses closed frontier models on visual grounding.
• Forged in Europe end-to-end and is deployable from Europe via our own Mistral Cloud infrastructure.
• Available to all via API today. Working with cybersecurity partners privately.
Open weights release end of October.
What others are saying
2 more-
Theo Otz @Totzenberger 2,101
A user compares model compute, claiming Mistral 4 reaches ~72% of rivals' intelligence using only 3,800 Grace Blackwell chips.
View on X -
Arthur Mensch says Mistral trained ML4 on 3,800 Grace Blackwell GPUs in its French datacenter.
View on X
Top replies on X
Sign in to see 5 top replies from X
From @anxuanng, @jeffacake117, @antipdoom and others. Spam removed, with English and Chinese translations.
Discussion 0
Sign up Sign in to join the discussion
No comments yet. Start the conversation.