10 Oct 2026
OllyGarden raises $4m seed to clean up OpenTelemetry data
Founded in 2025 by OpenTelemetry contributors, the company launched Minimum Viable Instrumentation with the round. Next Frontier Capital,…
Mistral AI's streaming speech-to-text model for Arabic dialects and Modern Standard Arabic, with open Apache 2.0 weights; about 4.4B parameters, built for live transcription with a configurable delay.
Not yet independently verified. The release date of 8 October 2026 rests only on the creation of the Hugging Face repository; Mistral AI published no announcement that we found, so the date may be earlier than the public release. The error rates are Mistral's own, from the model card, so no score is stored. No independent report was found. We will update this when it can be confirmed, and remove this note.
Voxtral Mini 4B Realtime Arabic is a streaming speech-to-text model from Mistral AI for Arabic dialects and Modern Standard Arabic. Its Hugging Face repository was created on 8 October 2026. It was built to cope with real-world audio in which speakers switch between Arabic and other languages.
The model transcribes 16 kHz audio as it arrives, and the delay between speech and text can be configured, which suits live captions and voice interfaces. It has about 4.4 billion parameters in BF16 and pairs a causal audio encoder with a language decoder. It was fine-tuned from Voxtral-Mini-4B-Realtime-2602, itself built on a 3B Ministral base.
Mistral reports an average character error rate of 8.82 per cent across seven Arabic benchmarks at a 480 millisecond delay, against 7.91 per cent for its Voxtral Transcribe Arabic. These are the maker's own figures, and the card warns that scores change with the normalisation method. Two new benchmarks, ISMA and Darija in the Wild, were developed with Morocco's digital transition ministry under a partnership signed in January 2026.
It suits developers who need live Arabic transcription they can host themselves, including for dialect-heavy and mixed-language speech. For other languages the card points to the original Voxtral Mini 4B Realtime. The model can be served with vLLM, streaming audio over a WebSocket at /v1/realtime, or run with Transformers version 5.2.0 or later.
| Benchmark | Official | Community avg |
|---|---|---|
| No benchmark scores yet. Be the first to add one. | ||
“Official” values are editor-approved and feed the ranking. “Community avg” is the mean of member submissions (shown for transparency; it never affects the ranking until an editor approves a value).
Sign in to add a benchmark score for this model.
An earned signal from verification, reviews, awards, transparency and engagement - the vendor can't buy it.
Updated 10/10/2026
Used it? Your experience helps other buyers decide.
Write a reviewNo questions yet. Be the first to ask about Voxtral Mini 4B Realtime Arabic.
Everything here links back to the same verified catalogue. Pick your next stop.