Mistral Large 4 enters Agent Arena's top 15 labs at -6.6% net improvement score
Overview
Arena (@arena) reports that Mistral Large 4, a preview model from Mistral AI, ranks #43 overall in its Agent Arena with a -6.6% net improvement score across more than 5,000 real-world agentic sessions.
Arena says this places Mistral among the top 15 labs, the only European lab there, and 11 places above its predecessor, Mistral Medium 3.5, which scored -12.60%.
Arena also says that Mistral Large 4's open weights are expected at the end of October, and that at its current score the model would rank #13 among open models. These are Arena's own figures and projections; the open-weights ranking is conditional on the release and on the score holding.
Written by AI from the articles below · updated Oct 9, 12:40 AM ET
Check the sources:
Article timeline
Follow the coverage from different perspectives. Times are ET.
- ArenaMistral Large 4 ranks in Agent Arena top 15 at -6.6% net score
AIMistral Large 4, a preview model from Mistral AI, ranks #43 overall in Agent Arena with a -6.6% net improvement score across more than 5,000 real-world agentic sessions. That is 11 rankings above its predecessor, Mistral Medium 3.5 (-12.60%), and places it in the top 15 labs, the only European lab there. Open weights are expected at the end of October, and at its current score the model would rank #13 among open models.
Heat trend
Not enough continuous observations to show a trend yet.