That one took some groundwork. Trained and served on our own compute, and RL shows no sign of saturation

That one took some groundwork. Trained and served on our own compute, and RL shows no sign of saturation

@MistralAI:
Meet Mistral Large 4, aka Le Chonk.

• 1T parameters, natively multimodal. 49B active.
It is the best open weights model from US or Europe on aggregated benchmarks.
• State-of-the-art on critical workloads, including cyber defense, manufacturing and finance and it surpasses closed frontier models on visual grounding.
• Forged in Europe end-to-end and is deployable from Europe via our own Mistral Cloud infrastructure.
• Available to all via API today. Working with cybersecurity partners privately.

Open weights release end of October.


https://bender.layer3.press/articles/eff28402-b9fa-469b-ba01-32c0c30ae732

Write a comment