Instructions to use AlicanKiraz0/Kizagan-TTS-v1.0 with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- VoxCPM
How to use AlicanKiraz0/Kizagan-TTS-v1.0 with VoxCPM:
import soundfile as sf from voxcpm import VoxCPM model = VoxCPM.from_pretrained("AlicanKiraz0/Kizagan-TTS-v1.0") wav = model.generate( text="VoxCPM is an innovative end-to-end TTS model from ModelBest, designed to generate highly expressive speech.", prompt_wav_path=None, # optional: path to a prompt speech for voice cloning prompt_text=None, # optional: reference text cfg_value=2.0, # LM guidance on LocDiT, higher for better adherence to the prompt, but maybe worse inference_timesteps=10, # LocDiT inference timesteps, higher for better result, lower for fast speed normalize=True, # enable external TN tool denoise=True, # enable external Denoise tool retry_badcase=True, # enable retrying mode for some bad cases (unstoppable) retry_badcase_max_times=3, # maximum retrying times retry_badcase_ratio_threshold=6.0, # maximum length restriction for bad case detection (simple but effective), it could be adjusted for slow pace speech ) sf.write("output.wav", wav, 16000) print("saved: output.wav") - Notebooks
- Google Colab
- Kaggle
Update README.md
Browse files
README.md
CHANGED
|
@@ -18,6 +18,8 @@ tags:
|
|
| 18 |
|
| 19 |
# Kahya-TTS-v1.0
|
| 20 |
|
|
|
|
|
|
|
| 21 |
Kahya is a Turkish text-to-speech adaptation of [OpenBMB's VoxCPM2](https://huggingface.co/openbmb/VoxCPM2), released by Alican Kiraz. It produces mono **48 kHz** speech and includes a reference-conditioned inference script for explicit sentence-by-sentence generation.
|
| 22 |
|
| 23 |
**v1.0 packages the already evaluated, merged Step 1500 checkpoint.** No additional fine-tuning or weight changes were performed for this release. The release adds a documented inference recipe, runnable commands, and source-derived evaluation results.
|
|
|
|
| 18 |
|
| 19 |
# Kahya-TTS-v1.0
|
| 20 |
|
| 21 |
+

|
| 22 |
+
|
| 23 |
Kahya is a Turkish text-to-speech adaptation of [OpenBMB's VoxCPM2](https://huggingface.co/openbmb/VoxCPM2), released by Alican Kiraz. It produces mono **48 kHz** speech and includes a reference-conditioned inference script for explicit sentence-by-sentence generation.
|
| 24 |
|
| 25 |
**v1.0 packages the already evaluated, merged Step 1500 checkpoint.** No additional fine-tuning or weight changes were performed for this release. The release adds a documented inference recipe, runnable commands, and source-derived evaluation results.
|