PAddle PARAllel text-to-speech toolKIT (supporting Tacotron2, Transformer TTS, FastSpeech2/FastPitch, SpeedySpeech, WaveFlow and Parallel WaveGAN)
Parakeet aims to provide a flexible, efficient and state-of-the-art text-to-speech toolkit for the open-source community. It is built on PaddlePaddle dynamic graph and includes many influential TTS models.
In order to facilitate exploiting the existing TTS models directly and developing the new ones, Parakeet selects typical models and provides their reference implementations in PaddlePaddle. Further more, Parakeet abstracts the TTS pipeline and standardizes the procedure of data preprocessing, common modules sharing, model configuration, and the process of training and synthesis. The models supported here include Text FrontEnd, end-to-end Acoustic models and Vocoders:
It's difficult to install some dependent libraries for this repo in Windows system, we recommend that you DO NOT use Windows system, please use
Make sure the library
libsndfile1is installed, e.g., on Ubuntu.
sudo apt-get install libsndfile1
See install for more details. This repo requires PaddlePaddle 2.1.2 or above.
git clone https://github.com/PaddlePaddle/Parakeet cd Parakeet pip install -e .
If some python dependent packages cannot be installed successfully, you can run the following script first. (replace
python3.6with your own python version)
bash sudo apt install -y python3.6-dev
See install for more details.
Entries to the introduction, and the launch of training and synthsis for different example models:
Check our website for audio sampels.
Parakeet is provided under the Apache-2.0 license.