Mistral releases a new open source model for speech generation
The model, which lets enterprises build voice agents for sales and customer engagement, puts Mistral in direct competition with the likes of ElevenLabs, Deepgram, and OpenAI.
The model, which lets enterprises build voice agents for sales and customer engagement, puts Mistral in direct competition with the likes of ElevenLabs, Deepgram, and OpenAI.
Relatively light at just 2 billion parameters, the model is meant for use with consumer-grade GPUs for those who want to self-host it. It currently supports 14 languages.
Speechify just launched a native Windows app that employs locally stored models to enable dictation and transcription across apps.
The Series B is the company’s second fundraise since it last raised capital in 2020. In that time, it has increased its ARR by 10x to $60 million.
Google’s new offline-first dictation app uses Gemma AI models to take on the apps like Wispr Flow.
DeepL says its tech could be used for real-time translation with meeting tools like Zoom and Microsoft Teams.
AI-powered dictation apps are useful for replying to emails, taking notes, and even coding through your voice
ElevenLabs reveals new investors, hits $500M ARR, and expands enterprise footprint as voice AI becomes a critical interface.
Ethos says it is onboarding 35,000 experts per week.
Wispr Flow says growth accelerated in India after its Hinglish rollout, even as voice AI products continue to face challenges.