3.5 KiB

Raw Blame History

Parakeet

Parakeet aims to provide a flexible, efficient, and state-of-the-art text-to-speech toolkit for the open-source community. It is built on PaddlePaddle dynamic graph and includes many influential TTS models.

Overview

To facilitate exploiting the existing TTS models directly and developing the new ones, Parakeet selects typical models and provides their reference implementations in PaddlePaddle. Furthermore, Parakeet abstracts the TTS pipeline and standardizes the procedure of data preprocessing, common modules sharing, model configuration, and the process of training and synthesis. The models supported here include Text FrontEnd, end-to-end Acoustic models, and Vocoders:

Audio samples

Check our website for audio sampels.

Released Model

Acoustic Model

FastSpeech2/FastPitch

SpeedySpeech

speedyspeech_nosil_baker_ckpt_0.5.zip

TransformerTTS

transformer_tts_ljspeech_ckpt_0.4.zip

Tacotron2

Vocoder

WaveFlow

waveflow_ljspeech_ckpt_0.3.zip

Parallel WaveGAN

Voice Cloning

Tacotron2_AISHELL3

tacotron2_aishell3_ckpt_0.3.zip

GE2E

ge2e_ckpt_0.3.zip

3.5 KiB Raw Blame History