Nari Labs open-sources low-latency, high-accuracy Qwen3 speech engine
Nari Labs open-sourced an inference engine for Qwen3-TTS and Qwen3-ASR, claiming sub-50 ms latency at 10 RPS and top accuracy/cost performance on Coval benchmarks versus closed services; the team published the repository and reported outperforming some official endpoints.