Qwen3-TTS (offline, 1.7B)

VoicesOfflineGPU optionalVoice cloningCommercial use OK

Multilingual neural TTS with 9 premium timbres + 3-second voice cloning. Heavy: ~3.5 GB model, 8+ GB RAM, CPU or CUDA.

Install it from Subtitld

Settings Add-ons Qwen3-TTS (offline, 1.7B) Install

Subtitld downloads it, checks its SHA-256, and keeps it updated.

Source on GitHub
Languages
10
Models
3
Voices
10
Add-on download
317–456 MB
RAM, at least
8 GB

Models

Model Download
qwen3-tts-1_7b-customvoice
3.4 GB
qwen3-tts-1_7b-base
3.4 GB
qwen3-tts-tokenizer-12hz
250 MB

Voices

10 voices in 4 languages.

  • en 2
  • ja 1
  • ko 1
  • zh-cn 5

Options

In Settings › Add-ons › Qwen3-TTS (offline, 1.7B) › Configure.

Inference device
Default: cpu
Weight precision
Default: auto
Default voice reference (3+ seconds, mono WAV recommended)
Default: empty

Used by the 'qwen3-clone' voice. Per-speaker overrides set in the speaker panel take precedence.

Default voice reference transcript
Default: empty

Optional. If provided alongside the reference clip, clone quality is noticeably better. Leave empty to use x-vector-only mode.

Skip auto-extracted reference transcript
Default: off

When the dubbing UI extracts a reference clip from the project audio, it normally also forwards the matching subtitle text as the transcript to improve clone accent fidelity. Enable this to suppress the transcript — qwen3 will fall back to its built-in ASR. Useful when the subtitle text paraphrases rather than transcribes verbatim, or when the source language is uncertain.

Languages

  • zh-cn
  • en
  • ja
  • ko
  • de
  • fr
  • ru
  • pt
  • es
  • it

Platforms

System Download
Linux x86-64 455.7 MB
macOS (universal) 345.8 MB
Windows x86-64 316.5 MB

License

Add-on
Apache-2.0 (wrapper) + Apache-2.0 (Qwen3-TTS weights)

Apache-2.0 (wrapper) + Apache-2.0 (Qwen3-TTS weights)

Versions

  • 1.0.6

    Automated release for qwen3-tts v1.0.6.

Other Voices add-ons