Mistral releases open-weight Voxtral Mini 4B Realtime 2602 speech model
mistralai/Voxtral-Mini-4B-Realtime-2602
AISummary
Mistral AI released Voxtral Mini 4B Realtime 2602, a multilingual realtime speech-transcription model with 13 supported languages under the Apache 2.0 license. The model has a configurable transcription delay from 240ms to 2.4s, and it matches leading offline open-source models at a 480ms delay. The source says it is optimized for on-device deployment and is currently supported only in vLLM.
AIWhy it matters
The source specifies the 480ms delay operating point, 4B size, Apache 2.0 license, and vLLM serving path, which matter for teams weighing realtime transcription deployment.
Source: Mistral AI · new models on Hugging Face · huggingface.co