Didispeech: A Large Scale Mandarin Speech Corpus

Tingwei Guo,Cheng Wen,Dongwei Jiang,Ne Luo,Ruixiong Zhang,Shuaijiang Zhao,Wubo Li,Cheng Gong,Wei Zou,Kun Han,Xiangang Li

Didispeech: A Large Scale Mandarin Speech Corpus

2021

Tingwei Guo
Cheng Wen
Dongwei Jiang
Ne Luo
Ruixiong Zhang
Shuaijiang Zhao
Wubo Li
Cheng Gong
Wei Zou
Kun Han
Xiangang Li

This paper introduces a new open-sourced Mandarin speech corpus, called DiDiSpeech. It consists of about 800 hours of speech data at 48kHz sampling rate from 6000 speakers and the corresponding texts. All speech data in the corpus is recorded in quiet environment and is suitable for various speech processing tasks, such as voice conversion, multi-speaker text-to-speech and automatic speech recognition. We conduct experiments with multiple speech tasks and evaluate the performance, showing that it is promising to use the corpus for both academic research and practical application. The corpus is available at this https URL.

Keywords:

Speech recognition
Speech processing
Speech corpus
Mandarin Chinese
QUIET
scale
Computer science

Correction
Source
Cite
Save
Machine Reading By IdeaReader

References

Citations