Mistral AI has unveiled Voxtral, its speech transcription model built around near-real-time processing speed. The announcement, framed as a research release, positions Voxtral as a competitive alternative in the automatic speech recognition (ASR) space. The "speed of sound" framing suggests the model's key differentiator is low-latency, fast transcription suitable for demanding production workloads.
Mistral AI has announced Voxtral, its debut audio-native language model family targeting speech recognition, multilingual transcription, and audio comprehension. Available in two sizes via Mistral's La Plateforme API, it extends the company's portfolio decisively into multimodal AI. The release positions Mistral as a full-stack AI provider capable of handling voice and audio alongside its established text and code capabilities.
ElevenLabs published a blog post titled “Introducing The Eleven Album.” Based only on the title, this appears to be a release or introductory announcement for a new project. The provided content does not include details on format, availability, technical features, collaborators, licensing, or whether it is directly tied to ElevenLabs’ voice AI tools.
ElevenLabs has introduced a new music generation model focused on finer-grained song editing. According to TechCrunch, users will be able to regenerate a section of a track without affecting the rest of the song. The headline also highlights genre switching mid-track, suggesting the model is aimed at more flexible AI music creation workflows.