Nari Labs achieves sub-50 ms TTS latency with Qwen3-TTS

Nari Labs' Qwen3-TTS 1.7B CustomVoice implementation hits sub-50 ms p95 time-to-first-audio at 10 RPS on a single H100, costing ~$2 per 1M characters vs ElevenLabs V3's $100/1M. It's the only one of five engines to achieve sub-50 ms p95 TTFA.
1 source
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- OpenAI reduces GPT-5.6 Sol API prices by over 20%
- Ben Thompson analysis questions US AI global dominance
- MiniMax M3 model gains support for SambaNova AI hardware
- Spline V2 rebuilds 3D editor, opens it to Claude Code via MCP
- Claude Code 2.1.239 adds residency cost premium