Qwen3-TTS · Issues· 55 open
Open on GitHubLocally synced open issues (discussions stay on GitHub)
- #353
[Commercial use clarification] May outputs from preset voice Sohee be used in monetized YouTube videos?
inactiveUpdated Sep 16, 2026 - #371
Finetuning loss misaligned with generation (double shift + leaked sub-talker hidden) — please merge #278
Updated Sep 16, 2026 - #294
Any plan to release 25Hz tts model ?
inactiveUpdated Sep 14, 2026 - #343
Fine-tuned 1.7B voice: short/opening utterances unstable (wrong timbre, occasional gender flip) while long-form is stable
inactiveUpdated Sep 13, 2026 - #369
25Hz tokenizer: qkv_attention_manual never masks padded keys (bool masked_fill), short attention windows get wrong outputs
Updated Sep 13, 2026 - #350
Multi-concurrency inference not supported
inactiveUpdated Sep 7, 2026 - #347
Verify evals on Papers with Code
inactiveUpdated Sep 6, 2026 - #365
finetuning/sft_12hz.py: checkpoint save is non-atomic (can silently leave base weights) and peaks at 2x model size in host RAM
Updated Sep 4, 2026 - #323
Fine-tuning Qwen3-TTS-12Hz-1.7B-Base for Bengali: Foreign/Hindi Accent Bias and Lack of Naturalness
inactiveUpdated Aug 26, 2026 - #340
AMD MI300X - Qwen/Qwen3-TTS-12Hz-1.7B-Base - Voice Cloning Segmentation fault (core dumped)
inactiveUpdated Aug 22, 2026 - #341
Voice-clone (ICL) generation echoes the tail of the reference audio before the requested text
Updated Aug 20, 2026 - #359
Feature request: native Italian preset voices (male + female)
Updated Aug 19, 2026 - #179
Finetuning Base results in progressively faster speech with every epoch
Updated Aug 19, 2026 - #337
The talker_hidden_states not obtained correctly
Updated Aug 18, 2026 - #55
The audio generated for Japanese and Korean is incomplete.
Updated Aug 17, 2026