ASR-project/Multilingual-PR
Phoneme Recognition using pre-trained models Wav2vec2, HuBERT and WavLM. Throughout this project, we compared specifically three different self-supervised models, Wav2vec (2019, 2020), HuBERT (2021) and WavLM (2022) pretrained on a corpus of English speech that we will use in various ways to perform phoneme recognition for different languages with a network trained with Connectionist Temporal Classification (CTC) algorithm.
⭐ 265
⑂ 22
Python
· 2022-05-10推送
265
Watchers
0
贡献者
0
Commits
0
Releases
2
Open Issues
2022-05-10
最近推送
原文
中文
暂无 README