Mistral Large 4 trained on own compute, RL shows no saturation
Original titleThat one took some groundwork. Trained and served on our own compute, and RL shows no sign of saturation
AISummary
Arthur Mensch says Mistral trained its model on its own compute, and reinforcement learning shows no sign of saturating. The post accompanies Mistral's announcement of Mistral Large 4, a 1T-parameter natively multimodal model with 49B active parameters, available via API today and with open weights planned for end of October.
Source: Arthur Mensch · x.comPublished · added here