Instructions to use mehdi-hf/pocket-tts-farsi with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Pocket-TTS
How to use mehdi-hf/pocket-tts-farsi with Pocket-TTS:
from pocket_tts import TTSModel import scipy.io.wavfile tts_model = TTSModel.load_model("mehdi-hf/pocket-tts-farsi") voice_state = tts_model.get_state_for_audio_prompt( "hf://kyutai/tts-voices/alba-mackenna/casual.wav" ) audio = tts_model.generate_audio(voice_state, "Hello world, this is a test.") # Audio is a 1D torch tensor containing PCM data. scipy.io.wavfile.write("output.wav", tts_model.sample_rate, audio.numpy()) - Notebooks
- Google Colab
- Kaggle
finetune guide
#2
by devops724 - opened
Hi there,
thanks for share this great model
is there any finetune guide with example code and dataset
best regards
Hi There,
Yes, all the coding including training, preprocessing, evaluations, etc are on the github of this project: https://github.com/mallahyari/pocket-tts
-Mehdi