Nari Labs' Qwen3-TTS and Qwen3-ASR models rank at the top of Coval's voice AI benchmarks, achieving leading latency and accuracy metrics while offering competitive pricing. The Qwen3-ASR Fast model achieves 44ms latency and 3.6% word error rate for speech-to-text, while Qwen3-TTS Fast delivers 63ms latency and 3.8% WER for text-to-speech, both priced significantly lower than competitors like ElevenLabs and AssemblyAI.
Nari launches Qwen3-ASR, a speech-to-text model ranked in the top 5 for accuracy on Coval Benchmarks with sub-50ms latency and the lowest pricing at $0.06 per hour. The company entered public beta today with free limited-time access to its Fast and Standard endpoints for real-time transcription.