Google has launched Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS, new speech generation models supporting over 100 languages. The models ranked first and second respectively on Hume AI's Overall Quality Index, with Gemini 3.8 Flash TTS also topping Hume AI's Voice Design Benchmark and its accent modeling category. Google says the models improve on long-form content and dual-speaker screenplay control compared to the prior Gemini 3.1 Flash TTS.
blog.google
· 2026-09-23
Nari Labs announced its Qwen3-TTS and Qwen3-ASR models achieved leading positions on Coval's voice AI benchmark, which measures latency and word error rate for speech AI systems. The ASR Fast model hit 44ms latency with 3.6% error rate at $0.12/hour, while the TTS Fast model achieved 63ms latency with a best-in-class 3.8% error rate, undercutting competitors like AssemblyAI and Deepgram on price.
narilabs.com
· 2026-09-14
Indian audio storytelling platform Pocket FM has doubled its annualized revenue run rate to $500 million over the past year, driven largely by AI-generated content. CEO Rohan Nayak said AI now accounts for 93% of the platform's overall catalog and 99% of new releases, though humans still shape the underlying story ideas.
techcrunch.com
· 2026-09-10