在抑郁自动检测的情况下，阅读与自发言语的比较

2017 8th IEEE International Conference on Cognitive Infocommunications (CogInfoCom) Pub Date : 2017-09-01 DOI:10.1109/COGINFOCOM.2017.8268245

G. Kiss, K. Vicsi

{"title":"在抑郁自动检测的情况下，阅读与自发言语的比较","authors":"G. Kiss, K. Vicsi","doi":"10.1109/COGINFOCOM.2017.8268245","DOIUrl":null,"url":null,"abstract":"In this paper, read and spontaneous speech have been compared in the light of automatic depression detection by speech processing. First, statistical analysis was carried out to select those acoustic features that differ significantly between healthy and depressed subjects in case of these two types of speech, separately for both gender. Secondly, statistical examination and classification experiments were prepared to compare the values of the selected features for the two types of speech. We were looking for the answer to which type of speech can be used to achieve better automatic depression detection results. As it was expected, the tempo related features, such as articulation rate, speech rate, and pause lengths are useful in case of spontaneous speech, while formants trajectories can be used only in case of read speech, because their values are mainly influenced by the linguistic content of the speech. Despite the significant differences of the features' values between read and spontaneous speech, there were no major differences in the detection accuracies. 83% detection accuracy was archived with read speech samples, and 86%detection accuracy was achieved with spontaneous speech samples.","PeriodicalId":212559,"journal":{"name":"2017 8th IEEE International Conference on Cognitive Infocommunications (CogInfoCom)","volume":"40 1","pages":"0"},"PeriodicalIF":0.0000,"publicationDate":"2017-09-01","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"17","resultStr":"{\"title\":\"Comparison of read and spontaneous speech in case of automatic detection of depression\",\"authors\":\"G. Kiss, K. Vicsi\",\"doi\":\"10.1109/COGINFOCOM.2017.8268245\",\"DOIUrl\":null,\"url\":null,\"abstract\":\"In this paper, read and spontaneous speech have been compared in the light of automatic depression detection by speech processing. First, statistical analysis was carried out to select those acoustic features that differ significantly between healthy and depressed subjects in case of these two types of speech, separately for both gender. Secondly, statistical examination and classification experiments were prepared to compare the values of the selected features for the two types of speech. We were looking for the answer to which type of speech can be used to achieve better automatic depression detection results. As it was expected, the tempo related features, such as articulation rate, speech rate, and pause lengths are useful in case of spontaneous speech, while formants trajectories can be used only in case of read speech, because their values are mainly influenced by the linguistic content of the speech. Despite the significant differences of the features' values between read and spontaneous speech, there were no major differences in the detection accuracies. 83% detection accuracy was archived with read speech samples, and 86%detection accuracy was achieved with spontaneous speech samples.\",\"PeriodicalId\":212559,\"journal\":{\"name\":\"2017 8th IEEE International Conference on Cognitive Infocommunications (CogInfoCom)\",\"volume\":\"40 1\",\"pages\":\"0\"},\"PeriodicalIF\":0.0000,\"publicationDate\":\"2017-09-01\",\"publicationTypes\":\"Journal Article\",\"fieldsOfStudy\":null,\"isOpenAccess\":false,\"openAccessPdf\":\"\",\"citationCount\":\"17\",\"resultStr\":null,\"platform\":\"Semanticscholar\",\"paperid\":null,\"PeriodicalName\":\"2017 8th IEEE International Conference on Cognitive Infocommunications (CogInfoCom)\",\"FirstCategoryId\":\"1085\",\"ListUrlMain\":\"https://doi.org/10.1109/COGINFOCOM.2017.8268245\",\"RegionNum\":0,\"RegionCategory\":null,\"ArticlePicture\":[],\"TitleCN\":null,\"AbstractTextCN\":null,\"PMCID\":null,\"EPubDate\":\"\",\"PubModel\":\"\",\"JCR\":\"\",\"JCRName\":\"\",\"Score\":null,\"Total\":0}","platform":"Semanticscholar","paperid":null,"PeriodicalName":"2017 8th IEEE International Conference on Cognitive Infocommunications (CogInfoCom)","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.1109/COGINFOCOM.2017.8268245","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"","JCRName":"","Score":null,"Total":0}

引用次数: 17

摘要

本文从语音处理自动抑郁检测的角度，对阅读语音和自发语音进行了比较。首先，对健康受试者和抑郁受试者在这两种类型的言语情况下，分别进行统计分析，选择具有显著差异的声学特征。其次，进行统计检验和分类实验，比较两类语音所选择的特征值。我们正在寻找答案，哪种类型的语音可以达到更好的自动抑郁检测结果。正如预期的那样，与节奏相关的特征，如发音率、言语率和停顿长度，在自发语音的情况下是有用的，而共振子轨迹只能在阅读语音的情况下使用，因为它们的值主要受语音的语言内容的影响。尽管阅读语音和自发语音的特征值存在显著差异，但检测准确率没有显著差异。对读语音样本的检测准确率达到83%，对自发语音样本的检测准确率达到86%。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

查看原文

微信好友朋友圈 QQ好友复制链接

本刊更多论文

Comparison of read and spontaneous speech in case of automatic detection of depression

In this paper, read and spontaneous speech have been compared in the light of automatic depression detection by speech processing. First, statistical analysis was carried out to select those acoustic features that differ significantly between healthy and depressed subjects in case of these two types of speech, separately for both gender. Secondly, statistical examination and classification experiments were prepared to compare the values of the selected features for the two types of speech. We were looking for the answer to which type of speech can be used to achieve better automatic depression detection results. As it was expected, the tempo related features, such as articulation rate, speech rate, and pause lengths are useful in case of spontaneous speech, while formants trajectories can be used only in case of read speech, because their values are mainly influenced by the linguistic content of the speech. Despite the significant differences of the features' values between read and spontaneous speech, there were no major differences in the detection accuracies. 83% detection accuracy was archived with read speech samples, and 86%detection accuracy was achieved with spontaneous speech samples.

求助全文

通过发布文献求助，成功后即可免费获取论文全文。去求助

来源期刊

2017 8th IEEE International Conference on Cognitive Infocommunications (CogInfoCom)

自引率

0.00%

发文量

期刊最新文献

Classification of cognitive load using voice features: A preliminary investigation A case study on time-interval fuzzy cognitive maps in a complex organization Á bilingual comparison of MaxEnt-and RNN-based punctuation restoration in speech transcripts Presentation of a medieval church in MaxWhere Blocklino: A graphical language for Arduino