# Voxtral transcribes at the speed of sound.

DevFeed: [Voxtral transcribes at the speed of sound.](<https://devfeed.tech/articles/voxtral-transcribes-at-the-speed-of-sound-7134.md>)

Original publisher: [Read original article](<https://mistral.ai/news/voxtral-transcribe-2/>)

Published: 2026-02-04T16:00:00Z

Content type: article

Language: en

Sources: [Mistral AI Blog](<https://devfeed.tech/sources/mistral-ai-blog.md>)

Topics: [asr](<https://devfeed.tech/topics/asr.md>), [voice ai](<https://devfeed.tech/topics/voice-ai.md>), [Low Latency](<https://devfeed.tech/topics/low-latency.md>), [real-time](<https://devfeed.tech/topics/real-time.md>), [Streaming](<https://devfeed.tech/topics/streaming.md>), [Open Source Models & Datasets](<https://devfeed.tech/topics/open-source-models-datasets.md>), [Security](<https://devfeed.tech/topics/security.md>), [Benchmark](<https://devfeed.tech/topics/benchmark.md>)

Tags: [agents](<https://devfeed.tech/tags/agents.md>), [apache](<https://devfeed.tech/tags/apache.md>), [arabic](<https://devfeed.tech/tags/arabic.md>), [audio](<https://devfeed.tech/tags/audio.md>), [batch](<https://devfeed.tech/tags/batch.md>), [benchmark](<https://devfeed.tech/tags/benchmark.md>), [benchmarks](<https://devfeed.tech/tags/benchmarks.md>), [cost](<https://devfeed.tech/tags/cost.md>), [efficiency](<https://devfeed.tech/tags/efficiency.md>), [low-latency](<https://devfeed.tech/tags/low-latency.md>), [models](<https://devfeed.tech/tags/models.md>), [open](<https://devfeed.tech/tags/open.md>), [performance](<https://devfeed.tech/tags/performance.md>), [security](<https://devfeed.tech/tags/security.md>), [speech](<https://devfeed.tech/tags/speech.md>), [streaming](<https://devfeed.tech/tags/streaming.md>), [transcription](<https://devfeed.tech/tags/transcription.md>)

## AI overview

Mistral introduces Voxtral Transcribe 2, a family of speech-to-text models comprising Voxtral Mini Transcribe V2 for batch transcription and Voxtral Realtime for live applications. The release highlights speaker diarization, word-level timestamps, multilingual transcription in 13 languages, configurable sub-200 ms latency, streaming transcription, and open Apache 2.0 weights for edge deployment.

## Source excerpt

The most powerful AI platform for enterprises. Customize, fine-tune, and deploy AI assistants, autonomous agents, and multimodal AI with open models.