PAddle PARAllel text-to-speech toolKIT (supporting WaveFlow, WaveNet, Transformer TTS and Tacotron2)
Parakeet aims to provide a flexible, efficient and state-of-the-art text-to-speech toolkit for the open-source community. It is built on PaddlePaddle Fluid dynamic graph and includes many influential TTS models proposed by Baidu Research and other research groups.
In particular, it features the latest WaveFlow model proposed by Baidu Research.
In order to facilitate exploiting the existing TTS models directly and developing the new ones, Parakeet selects typical models and provides their reference implementations in PaddlePaddle. Further more, Parakeet abstracts the TTS pipeline and standardizes the procedure of data preprocessing, common modules sharing, model configuration, and the process of training and synthesis. The models supported here include Vocoders and end-to-end TTS models:
And more will be added in the future.
Make sure the library
libsndfile1is installed, e.g., on Ubuntu.
sudo apt-get install libsndfile1
See install for more details. This repo requires PaddlePaddle 2.0.0rc1 or above.
pip install -U paddle-parakeet
bash git clone https://github.com/PaddlePaddle/Parakeet cd Parakeet pip install -e .
See install for more details.
Entries to the introduction, and the launch of training and synthsis for different example models:
Check our website for audio sampels.
Models pretrained on LJSpeech can be downloaded here.
Parakeet is provided under the Apache-2.0 license.