r/LocalLLaMA
· Communities
I built an Android app for Qwen3 TTS on-device
It's more of a showcase demo at the moment, but Qwen3TTS 0.6B runs usable fast in Q4. On my Galaxy S25 I get around 0.5 realtime speed on CPU, depending of the length of the text. Short text is faster, longer text slower. The app uses a custom GGML based C++ backend (https://github.com/Danmoreng/qwen3-tts.cpp). Sadly I