🇫🇷 Mistral Large 4 is a 1T-parameter MoE trained on just 4,000 Blackwell GPUs
Nicknamed "Le Chonk." 49B active, natively multimodal, API preview live now, open weights Oct 27. That's roughly 2-3x less compute than the big Chinese labs, per Mistral.
It's strong on legal, finance, and visual grounding. Weaker than rivals on agentic coding, which is the benchmark everyone actually watches.
It's an efficiency flex.
Post #662
628