PSYCH OpenIR  > 中国科学院行为科学重点实验室
CDSD: Chinese Dysarthria Speech Database
Sun, Mengyi1,2; Gao, Ming2; Kang, Xinchen3; Wang, Shiru1,2; Du, Jun3; Yao, Dengfeng4,4,5; Wang, Su-Jing1,2
通讯作者邮箱yao, dengfeng ; wang, su-jing
摘要

We present the Chinese Dysarthria Speech Database (CDSD)1 as a valuable resource for dysarthria research. This database comprises speech data from 24 participants with dysarthria. Among these participants, one recorded an additional 10 hours of speech data, while each recorded one hour, resulting in 34 hours of speech material. To accommodate participants with varying cognitive levels, our text pool primarily consists of content from the AISHELL-1 dataset and speeches by primary and secondary school students. When participants read these texts, they must use a mobile device or the ZOOM F8n multi-track field recorder to record their speeches. In this paper, we elucidate the data collection and annotation processes and present an approach for establishing a baseline for dysarthric speech recognition. Furthermore, we conducted a speaker-dependent dysarthric speech recognition experiment using an additional 10 hours of speech data from one of our participants. Our research findings indicate that, through extensive data-driven model training, fine-tuning limited quantities of specific individual data yields commendable results in speaker-dependent dysarthric speech recognition. However, we observe significant variations in recognition results among different dysarthric speakers. These insights provide valuable reference points for speaker-dependent dysarthric speech recognition.

2023
语种英语
DOI10.48550/arXiv.2310.15930
发表期刊arXiv
期刊论文类型实证研究
收录类别EI
引用统计
文献类型期刊论文
条目标识符http://ir.psych.ac.cn/handle/311026/46269
专题中国科学院行为科学重点实验室
作者单位1.CAS Key Laboratory of Behavioral Science, Institute of Psychology, 16, Lincui Road, Chaoyang District, Beijing; 100101, China
2.Department of Psychology, University of the Chinese Academy of Sciences, No.1 Yanqihu East Rd, Huairou District, Beijing; 101408, China
3.University of Science and Technology of China, No.96, JinZhai Road Baohe District, Anhui, Hefei; 230026, China
4.Beijing Key Laboratory of Information Service Engineering, Beijing Union University, No.97, Beisihuan East Road, Chaoyang District, Beijing; 100101, China
5.Laboratory of Computational Linguistics, School of Humanities, Tsinghua University, No.1, Tsinghua Garden, Haidian District, Beijing; 100084, China
6.Center for Psychology and Cognitive Sciences, Tsinghua University, No.1, Tsinghua Garden, Haidian District, Beijing; 100084, China
第一作者单位中国科学院行为科学重点实验室
推荐引用方式
GB/T 7714
Sun, Mengyi,Gao, Ming,Kang, Xinchen,et al. CDSD: Chinese Dysarthria Speech Database[J]. arXiv,2023.
APA Sun, Mengyi.,Gao, Ming.,Kang, Xinchen.,Wang, Shiru.,Du, Jun.,...&Wang, Su-Jing.(2023).CDSD: Chinese Dysarthria Speech Database.arXiv.
MLA Sun, Mengyi,et al."CDSD: Chinese Dysarthria Speech Database".arXiv (2023).
条目包含的文件
条目无相关文件。
个性服务
推荐该条目
保存到收藏夹
查看访问统计
导出为Endnote文件
谷歌学术
谷歌学术中相似的文章
[Sun, Mengyi]的文章
[Gao, Ming]的文章
[Kang, Xinchen]的文章
百度学术
百度学术中相似的文章
[Sun, Mengyi]的文章
[Gao, Ming]的文章
[Kang, Xinchen]的文章
必应学术
必应学术中相似的文章
[Sun, Mengyi]的文章
[Gao, Ming]的文章
[Kang, Xinchen]的文章
相关权益政策
暂无数据
收藏/分享
所有评论 (0)
暂无评论
 

除非特别说明,本系统中所有内容都受版权保护,并保留所有权利。