Easy-to-use Speech Toolkit including SOTA/Streaming ASR with punctuation, influential TTS with text frontend, Speaker Verification System and End-to-End Speech Simultaneous Translation.
You can not select more than 25 topics Topics must start with a letter or number, can include dashes ('-') and can be up to 35 characters long.
 
 
 
 
 
 
Go to file
Hui Zhang 8977f1850e
close editdistance package format warning
3 years ago
.github
.pre-commit-hooks
.travis
deepspeech close editdistance package format warning 3 years ago
docs test refactor collator 3 years ago
examples Merge pull request #872 from Jackwaterveg/Hub 3 years ago
hub add requirements for hub 3 years ago
speechnn fix for kaldi 3 years ago
tests bool logical, sum and multiply op; ctc grad norm; support old and new pd api 3 years ago
third_party Kaldi (#839) 3 years ago
tools fix sctk install 3 years ago
utils reader default type is mat, sound need explicitlyc specify 3 years ago
.bashrc fix conf and readme 3 years ago
.clang-format
.flake8 refactor feature, dict and argument for new config format 3 years ago
.gitconfig
.gitignore fix sctk install 3 years ago
.mergify.yml
.pre-commit-config.yaml
.style.yapf
.travis.yml
.vimrc
LICENSE
README.md fix doc link 3 years ago
env.sh Merge pull request #763 from PaddlePaddle/st 3 years ago
requirements.txt close editdistance package format warning 3 years ago
setup.sh

README.md

PaddlePaddle Speech to Any toolkit

License python version support os

DeepSpeech is an open-source implementation of end-to-end Automatic Speech Recognition engine, with PaddlePaddle platform. Our vision is to empower both industrial application and academic research on speech recognition, via an easy-to-use, efficient, samller and scalable implementation, including training, inference & testing module, and deployment.

Features

See feature list for more information.

Setup

All tested under:

  • Ubuntu 16.04
  • python>=3.7
  • paddlepaddle>=2.2.0rc

Please see install.

Getting Started

Please see Getting Started and tiny egs.

More Information

Questions and Help

You are welcome to submit questions in Github Discussions and bug reports in Github Issues. You are also welcome to contribute to this project.

License

DeepSpeech is provided under the Apache-2.0 License.

Acknowledgement

We depends on many open source repos. See References for more information.