Awesome
Parakeet has moved to PaddleSpeech, this repo will not update anymore, you can open issues of Parakeet in PaddleSpeech
Parakeet
Parakeet aims to provide a flexible, efficient and state-of-the-art text-to-speech toolkit for the open-source community. It is built on PaddlePaddle dynamic graph and includes many influential TTS models.
<div align="center"> <img src="docs/images/logo.png" width=300 /> <br> </div>News <img src="./docs/images/news_icon.png" width="40"/>
- Oct-12-2021, Refector examples code.
- Oct-12-2021, Parallel WaveGAN with LJSpeech. Check examples/GANVocoder/parallelwave_gan/ljspeech.
- Oct-12-2021, FastSpeech2/FastPitch with LJSpeech. Check examples/fastspeech2/ljspeech.
- Sep-14-2021, Reconstruction of TransformerTTS. Check examples/transformer_tts/ljspeech.
- Aug-31-2021, Chinese Text Frontend. Check examples/text_frontend.
- Aug-23-2021, FastSpeech2/FastPitch with AISHELL-3. Check examples/fastspeech2/aishell3.
- Aug-03-2021, FastSpeech2/FastPitch with CSMSC. Check examples/fastspeech2/baker.
- Jul-19-2021, SpeedySpeech with CSMSC. Check examples/speedyspeech/baker.
- Jul-01-2021, Parallel WaveGAN with CSMSC. Check examples/GANVocoder/parallelwave_gan/baker.
- Jul-01-2021, Montreal-Forced-Aligner. Check examples/use_mfa.
- May-07-2021, Voice Cloning in Chinese. Check examples/tacotron2_aishell3.
Overview
In order to facilitate exploiting the existing TTS models directly and developing the new ones, Parakeet selects typical models and provides their reference implementations in PaddlePaddle. Further more, Parakeet abstracts the TTS pipeline and standardizes the procedure of data preprocessing, common modules sharing, model configuration, and the process of training and synthesis. The models supported here include Text FrontEnd, end-to-end Acoustic models and Vocoders:
-
Text FrontEnd
- Rule based Chinese frontend.
-
Acoustic Models
-
Vocoders
-
Voice Cloning
Setup
It's difficult to install some dependent libraries for this repo in Windows system, we recommend that you DO NOT use Windows system, please use Linux
.
Make sure the library libsndfile1
is installed, e.g., on Ubuntu.
sudo apt-get install libsndfile1
Install PaddlePaddle
See install for more details. This repo requires PaddlePaddle 2.1.2 or above.
Install Parakeet
git clone https://github.com/PaddlePaddle/Parakeet
cd Parakeet
pip install -e .
If some python dependent packages cannot be installed successfully, you can run the following script first.
(replace python3.6
with your own python version)
sudo apt install -y python3.6-dev
See install for more details.
Examples
Entries to the introduction, and the launch of training and synthsis for different example models:
- >>> Chinese Text Frontend
- >>> FastSpeech2/FastPitch
- >>> Montreal-Forced-Aligner
- >>> Parallel WaveGAN
- >>> SpeedySpeech
- >>> Tacotron2_AISHELL3
- >>> GE2E
- >>> WaveFlow
- >>> TransformerTTS
- >>> Tacotron2
Audio samples
TTS models (Acoustic Model + Neural Vocoder)
Check our website for audio sampels.
Released Model
Acoustic Model
FastSpeech2/FastPitch
- fastspeech2_nosil_baker_ckpt_0.4.zip
- fastspeech2_nosil_aishell3_ckpt_0.4.zip
- fastspeech2_nosil_ljspeech_ckpt_0.5.zip
SpeedySpeech
TransformerTTS
Tacotron2
Vocoder
WaveFlow
Parallel WaveGAN
Voice Cloning
Tacotron2_AISHELL3
GE2E
License
Parakeet is provided under the Apache-2.0 license.