Show HN: Nari Qwen3-TTS and Qwen3-ASR – High accuracy, low latency and cost (narilabs.com)

7 points by toebee 2 hours ago

1 comment:

by asaiacai 6 minutes ago

This is really cool work! I'm curious like what do you see as the biggest lever for speeding up TTS models or from a technical perspective that this was a promising direction in the first place to push on. If I were to guess, some distillation but I'm certain there are probably TTS model aware architectural changes that just make inference wayyyy faster?

Data from: Hacker News, provided by Hacker News (unofficial) API