Easy-to-use Speech Toolkit including SOTA/Streaming ASR with punctuation, influential TTS with text frontend, Speaker Verification System and End-to-End Speech Simultaneous Translation.
You can not select more than 25 topics Topics must start with a letter or number, can include dashes ('-') and can be up to 35 characters long.
 
 
 
 
 
 
Go to file
chenfeiyu abccbb5c7d
WIP: add kaldi-style frame and stft
3 years ago
.github add stale config (#604) 3 years ago
.notebook add test notebook 3 years ago
.pre-commit-hooks add pre commit hooks 3 years ago
.travis E2E/Streaming Transformer/Conformer ASR (#578) 3 years ago
deepspeech Merge pull request #657 from PaddlePaddle/add_utt 3 years ago
doc remove useless doc 3 years ago
examples Merge pull request #657 from PaddlePaddle/add_utt 3 years ago
tests remove sequnce_mask and change ds2 export audio shape to [B,T,D] (#639) 3 years ago
third_party WIP: add kaldi-style frame and stft 3 years ago
tools downgrad python env version 3 years ago
utils refactor g2p egs (#630) 3 years ago
.clang-format E2E/Streaming Transformer/Conformer ASR (#578) 3 years ago
.flake8 E2E/Streaming Transformer/Conformer ASR (#578) 3 years ago
.gitconfig E2E/Streaming Transformer/Conformer ASR (#578) 3 years ago
.gitignore more decoding method (#618) 3 years ago
.mergify.yml ctc decoding weight 0.5 (#614) 3 years ago
.pre-commit-config.yaml E2E/Streaming Transformer/Conformer ASR (#578) 3 years ago
.style.yapf Add ci and code format checking. 7 years ago
.travis.yml E2E/Streaming Transformer/Conformer ASR (#578) 3 years ago
.vimrc E2E/Streaming Transformer/Conformer ASR (#578) 3 years ago
LICENSE Create License 7 years ago
README.md remove sequnce_mask and change ds2 export audio shape to [B,T,D] (#639) 3 years ago
README_cn.md fix doc link (#627) 3 years ago
env.sh add tarball utils (#626) 3 years ago
requirements.txt remove sequnce_mask and change ds2 export audio shape to [B,T,D] (#639) 3 years ago
setup.sh chinese char/word ngram lm (#613) 3 years ago

README.md

中文版

PaddlePaddle ASR toolkit

License python version support os

PaddleASR is an open-source implementation of end-to-end Automatic Speech Recognition (ASR) engine, with PaddlePaddle platform. Our vision is to empower both industrial application and academic research on speech recognition, via an easy-to-use, efficient, samller and scalable implementation, including training, inference & testing module, and deployment.

Features

See feature list for more information.

Setup

  • python>=3.7
  • paddlepaddle>=2.1.0

Please see install.

Getting Started

Please see Getting Started and tiny egs.

More Information

Questions and Help

You are welcome to submit questions in Github Discussions and bug reports in Github Issues. You are also welcome to contribute to this project.

License

DeepASR is provided under the Apache-2.0 License.

Acknowledgement

We depends on many open source repos. See References for more information.