Loading repository data…
Loading repository data…
NeoKazuya / repository
Enhanced Qwen3-TTS voice cloning GUI with multi-reference samples, variation generation, and audio preprocessing.
A transparent discovery signal based on current public GitHub metadata.
This score does not audit code, security, maintainers, documentation quality, or suitability. Verify the repository and its current documentation before adoption.
Clone any voice in seconds. 100% local, runs on your GPU.
An enhanced interface for Qwen3-TTS with multi-reference cloning, variation generation, and audio preprocessing.
install.bat on Windows| Feature | Description |
|---|---|
| 🎤 Voice Clone | Clone voices from short audio (3+ seconds) |
| 🎭 Create Voice | Combine multiple samples with per-file transcripts |
| 👤 Custom Voice | 9 preset speakers with emotion control |
| ✨ Voice Design | Create voices from text descriptions |
| 💾 Save & Load | Keep voices as portable .pt files |
| ⚙️ Settings | Configure data folder, persists across updates |
GTX 10 series (Pascal)? Use v1.2.3 which includes CUDA 12.4.
install.bat # One-time setup
run.bat # Launch app
docker-compose up --build
First run downloads ~4GB of models.
Apache 2.0 - See LICENSE
Built on Qwen3-TTS by Alibaba Cloud.