7 Sources
[1]
Mistral releases Voxtral, its first open source AI audio model | TechCrunch
As AI systems become more capable, speech is fast becoming the default way we communicate with machines. French AI startup Mistral has jumped into the audio race with its first open model, aiming to challenge the dominance of walled-off corporate systems with open-weight alternatives. On Tuesday,
[2]
Mistral launches Voxtral speech recognition model
Mistral has released an open automatic speech recognition (ASR) software bundle called Voxtral in a bid to undercut rivals on price and quality. The biz claims that using ASR in production has required a trade-off - using open-source models with high error rates and limited semantic understanding
[3]
Mistral's Voxtral goes beyond transcription with summarization, speech-triggered functions
Want smarter insights in your inbox? Sign up for our weekly newsletters to get only what matters to enterprise AI, data, and security leaders. Subscribe Now Mistral released an open-sourced voice model today that could rival paid voice AI, such as those from ElevenLabs and Hume AI, which the
[4]
Mistral Unveils Voxtral, Its Open-Source Bet to Rival OpenAI and ElevenLabs | AIM
The French AI startup is taking the open-source route to tackle competitors in the speech understanding AI models space. French AI startup Mistral has released Voxtral, a new family of open-source speech understanding models designed to deliver production-ready transcription and semantic audio
[5]
Mistral Voxtral: Open-source AI audio arrives
Mistral introduced Voxtral, its initial open-source AI audio model family, on Tuesday, aiming to provide businesses with a production-ready speech intelligence solution. This release challenges existing corporate systems by offering an open-weight alternative for audio processing. The company
[6]
Mistral Releases Its First Open-Source Speech Generation Models
Mistral's new speech generation model can detect multiple languages Mistral released its first speech understanding models on Tuesday. Dubbed Voxtral, it is an open-source audio generation artificial intelligence (AI) model that not only turns text into speech but can also understand text to
[7]
What is Voxtral: Mistral's open Source AI Audio Model, Key Features Explained
Voxtral is free, fast, Apache 2.0 licensed, and outperforms Whisper and GPT-4o mini in benchmarks. In July 2025, Mistral AI unveiled Voxtral, a powerful new entry in the world of AI audio models. Unlike most competitors, Voxtral is fully open source and designed for deeper audio understanding,
Share
Copy Link
French AI startup Mistral releases Voxtral, an open-source speech recognition model family, aiming to provide affordable and accurate audio processing solutions for businesses while competing with established proprietary systems.
French AI startup Mistral has made a significant move in the artificial intelligence landscape with the release of Voxtral, its first family of open-source AI audio models
1
. This launch marks Mistral's entry into the competitive speech recognition market, challenging established proprietary systems with an open-weight alternative designed for business applications.
Source: Dataconomy
Voxtral aims to address a critical dilemma faced by developers in the field of speech recognition. Traditionally, they have had to choose between inexpensive open systems with limited accuracy and understanding, or well-functioning but closed systems that come with higher costs and less deployment flexibility
2
. Mistral positions Voxtral as a solution that offers "state-of-the-art accuracy and native semantic understanding in the open, at less than half the price of comparable APIs"3
.Voxtral is available in two main variants:
Additionally, Mistral offers Voxtral Mini Transcribe, a streamlined version of the 3B model specifically designed for transcription tasks
1
.The model can handle up to 32,000 tokens of context, allowing it to process approximately 30 minutes of audio for transcription or 40 minutes for comprehension
4
. Voxtral's capabilities extend beyond mere transcription, enabling users to ask questions about audio content, generate summaries, and even trigger real-time actions like API calls or function executions through voice commands1
.
Source: AIM
Voxtral boasts strong multilingual capabilities, supporting languages such as English, Spanish, French, Portuguese, Hindi, German, Dutch, and Italian for both transcription and comprehension
1
. Mistral claims that Voxtral outperforms existing models like OpenAI's Whisper, Gemini 2.5 Flash, and ElevenLabs' Scribe across various benchmarks, including FLEURS and Mozilla Common Voice4
.Mistral has made Voxtral accessible through multiple channels. Users can download the API from Hugging Face or test the models in Mistral's chatbot, Le Chat, free of charge. For those looking to integrate the API into their applications, pricing starts at a competitive rate of $0.001 per minute
5
. This pricing strategy positions Voxtral as a cost-effective alternative to existing solutions in the market.Related Stories
The release of Voxtral represents a significant development in the open-source AI community. By offering a high-performance, open-source alternative to proprietary speech recognition systems, Mistral is challenging the status quo and potentially democratizing access to advanced audio processing capabilities
3
.
Source: The Register
Mistral has indicated that it is actively expanding its audio team, with the goal of developing "near-human-like voice interfaces"
4
. This launch follows the recent introduction of Magistral, Mistral's reasoning-focused language model, demonstrating the company's commitment to innovation across various AI domains5
.As Mistral continues to make waves in the AI industry, reports suggest that the company is in talks to raise up to $1 billion in equity from investors, including Abu Dhabi's MGX fund
1
. This potential influx of capital could further accelerate Mistral's growth and development in the competitive AI landscape.Summarized by
Navi
[2]
[3]
[5]
26 Mar 2026•Technology

04 Feb 2026•Technology

08 May 2025•Technology

1
Technology

2
Policy and Regulation

3
Technology
