TianYuan
c0ee57d400
Merge branch 'develop' of https://github.com/PaddlePaddle/DeepSpeech into thchs30_MFA
3 years ago
TianYuan
7d4eff2b86
add MFA example for THCHS30
3 years ago
Hui Zhang
43b52082c3
Merge pull request #629 from PaddlePaddle/align
...
ctc alignment
3 years ago
Hui Zhang
4c9a1f6dc7
add align.sh and update run.sh
3 years ago
Hui Zhang
6ee67785f6
fix ctc alignment
3 years ago
Hui Zhang
717fe1e4bd
Merge pull request #680 from PaddlePaddle/checkpoint
...
checkpoint refactor to save disk space
3 years ago
Haoxin Ma
c0f7aac8fc
revise conf/*.yaml
3 years ago
Hui Zhang
8c0923b865
update gitignore; add gigaspeech
3 years ago
Hui Zhang
e106f243b4
dump dataset metadata
3 years ago
Hui Zhang
9e99f99b3c
add thchs30, aidatatang;
3 years ago
Hui Zhang
b8f190c12c
add thchs30 dataset
3 years ago
Hui Zhang
9b3acddd5d
fix conf for new datapipe; u2 export inputspec
3 years ago
Hui Zhang
f8d52e598f
Merge branch 'develop' into rsl
3 years ago
Hui Zhang
03e6952501
more detial of result
3 years ago
Hui Zhang
718bd30765
Merge pull request #688 from PaddlePaddle/ds2_fix
...
fix conf for ds2
3 years ago
Haoxin Ma
91e70a2857
multi gpus
3 years ago
Hui Zhang
019ae4b35c
fix conf for ds2
3 years ago
Hui Zhang
4b80b172d3
add model params
3 years ago
Hui Zhang
133a522fbb
ds2 default using 4gpu; new result of ds2
3 years ago
Haoxin Ma
8af2eb073a
revise config
3 years ago
Haoxin Ma
68bcc46940
save best and test on tiny/s0
3 years ago
Hui Zhang
68149cb9a7
fix config for new datapipeline
3 years ago
Hui Zhang
5a3a9e1f50
fix chunk default config; tarball ckpt prfix dir;
3 years ago
Hui Zhang
8c1bf1a730
fix ds2 conf for new data pipeline
3 years ago
Haoxin Ma
557427736e
move redundant params
3 years ago
Haoxin Ma
698d7a9bdb
move batch_size, work_nums, shuffle_method, sortagrad to collator
3 years ago
Haoxin Ma
6ee3033cc4
finish aishell/s0
3 years ago
Haoxin Ma
7bae32f384
revise example/ting/s1
3 years ago
Haoxin Ma
b9110af9d3
feat_dim, vocab_size
3 years ago
Blank
875139ca04
Merge branch 'develop' into spec_aug
3 years ago
Hui Zhang
1cd88d2619
Merge pull request #657 from PaddlePaddle/add_utt
...
add utt to datapipeline
3 years ago
Haoxin Ma
b4bda290aa
fix bugs
3 years ago
Hui Zhang
255a6f004e
result with specaug
3 years ago
Haoxin Ma
279348d786
move process utt to collator
3 years ago
Haoxin Ma
8781ab58cf
fix export and run.sh
3 years ago
Haoxin Ma
a58b1cb30a
add result output
3 years ago
Hui Zhang
9bb1806246
fix clip conf
3 years ago
Hui Zhang
b3d23e6af3
add stream conf
3 years ago
Haoxin Ma
c8368410e2
utt datapipeline
3 years ago
Hui Zhang
69dfc2a5fa
fix mask for bool type; fix other
3 years ago
Hui Zhang
34689bd1df
add crf
4 years ago
Hui Zhang
de780a0c8a
add libri ds2 exp result
4 years ago
chenfeiyu
4c3c5546e7
1. use space as separator;
...
2. add docstring for some functions.
4 years ago
Hui Zhang
759b2e0b2d
fix ds2 default config
4 years ago
Hui Zhang
b3bc451328
remove sequnce_mask and change ds2 export audio shape to [B,T,D] ( #639 )
...
* remove sequnce_mask
* format
* fix ds2 export audio shape from B,D,T to B,T,D
4 years ago
Hui Zhang
749a113037
add Cedict egs & pypinyin using jieba as wordseg & phkit using local pypinyin package ( #637 )
...
* down and parse cedict
* remove useless
* using third party python pinyin
* jieba as default wordseg
* remove useless
* remove pinyin dict
* using jieba.lcut
* remove doc of cedict egs
* add fan2jian test
* add description for say_digit
4 years ago
Hui Zhang
1b373bfcd7
refactor g2p egs ( #630 )
...
* refactor g2p egs
* add sha-bone; remove avg.sh from egs;
4 years ago
Feiyu Chan
075635d2b4
add a g2p example ( #623 )
...
* add an example to convert transcription to pinyin with pypinyin and jieba
* format code
* 1. remove script for data downloading, since Baker dataset is not easily downloaded via terminal;
2. remove pypinyin as an extra requirement; it is alreay required by the main project;
3. clean code.
* change output format
4 years ago
Hui Zhang
0a7958b3f1
add tarball utils ( #626 )
4 years ago
Hui Zhang
37c5324138
fix result; add feature list
4 years ago