Nonlinear vocal phenomena and speech intelligibility.

Andrey Anikin, David Reby, Katarzyna Pisanski
{"title":"Nonlinear vocal phenomena and speech intelligibility.","authors":"Andrey Anikin, David Reby, Katarzyna Pisanski","doi":"10.1098/rstb.2024.0254","DOIUrl":null,"url":null,"abstract":"<p><p>At some point in our evolutionary history, humans lost vocal membranes and air sacs, representing an unexpected simplification of the vocal apparatus relative to other great apes. One hypothesis is that these simplifications represent anatomical adaptations for speech because a simpler larynx provides a suitably stable and tonal vocal source with fewer nonlinear vocal phenomena (NLP). The key assumption that NLP reduce speech intelligibility is indirectly supported by studies of dysphonia, but it has not been experimentally tested. Here, we manipulate NLP in vocal stimuli ranging from single vowels to sentences, showing that the vocal source needs to be stable, but not necessarily tonal, for speech to be readily understood. When the task is to discriminate synthesized monophthong and diphthong vowels, continuous NLP (subharmonics, amplitude modulation and even deterministic chaos) actually improve vowel perception in high-pitched voices, likely because the resulting dense spectrum reveals formant transitions. Rough-sounding voices also remain highly intelligible when continuous NLP are added to recorded words and sentences. In contrast, voicing interruptions and pitch jumps dramatically reduce speech intelligibility, likely by interfering with voicing contrasts and normal intonation. We argue that NLP were not eliminated from the human vocal repertoire as we evolved for speech, but only brought under better control.This article is part of the theme issue 'Nonlinear phenomena in vertebrate vocalizations: mechanisms and communicative functions'.</p>","PeriodicalId":19872,"journal":{"name":"Philosophical Transactions of the Royal Society B: Biological Sciences","volume":"380 1923","pages":"20240254"},"PeriodicalIF":4.7000,"publicationDate":"2025-04-03","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"https://www.ncbi.nlm.nih.gov/pmc/articles/PMC11966171/pdf/","citationCount":"0","resultStr":null,"platform":"Semanticscholar","paperid":null,"PeriodicalName":"Philosophical Transactions of the Royal Society B: Biological Sciences","FirstCategoryId":"99","ListUrlMain":"https://doi.org/10.1098/rstb.2024.0254","RegionNum":2,"RegionCategory":"生物学","ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"Q1","JCRName":"BIOLOGY","Score":null,"Total":0}
引用次数: 0

Abstract

At some point in our evolutionary history, humans lost vocal membranes and air sacs, representing an unexpected simplification of the vocal apparatus relative to other great apes. One hypothesis is that these simplifications represent anatomical adaptations for speech because a simpler larynx provides a suitably stable and tonal vocal source with fewer nonlinear vocal phenomena (NLP). The key assumption that NLP reduce speech intelligibility is indirectly supported by studies of dysphonia, but it has not been experimentally tested. Here, we manipulate NLP in vocal stimuli ranging from single vowels to sentences, showing that the vocal source needs to be stable, but not necessarily tonal, for speech to be readily understood. When the task is to discriminate synthesized monophthong and diphthong vowels, continuous NLP (subharmonics, amplitude modulation and even deterministic chaos) actually improve vowel perception in high-pitched voices, likely because the resulting dense spectrum reveals formant transitions. Rough-sounding voices also remain highly intelligible when continuous NLP are added to recorded words and sentences. In contrast, voicing interruptions and pitch jumps dramatically reduce speech intelligibility, likely by interfering with voicing contrasts and normal intonation. We argue that NLP were not eliminated from the human vocal repertoire as we evolved for speech, but only brought under better control.This article is part of the theme issue 'Nonlinear phenomena in vertebrate vocalizations: mechanisms and communicative functions'.

Abstract Image

Abstract Image

Abstract Image

查看原文
分享 分享
微信好友 朋友圈 QQ好友 复制链接
本刊更多论文
非线性声音现象与语音可理解性。
在我们进化史上的某个时刻,人类失去了声膜和气囊,这代表了相对于其他类人猿的发声器官的意外简化。一种假设是,这些简化代表了语言的解剖学适应,因为更简单的喉部提供了一个适当的稳定和音调的发声源,减少了非线性发声现象(NLP)。语音障碍的研究间接支持了NLP降低语音清晰度的关键假设,但尚未经过实验验证。在这里,我们在从单个元音到句子的声音刺激中操纵NLP,表明声音来源需要稳定,但不一定是音调,以便易于理解。当任务是区分合成的单元音和双元音时,连续的NLP(亚谐波、调幅甚至确定性混沌)实际上提高了对高音语音的元音感知,可能是因为由此产生的密集频谱揭示了形成峰转换。当连续的自然语言处理添加到录制的单词和句子中时,粗糙的声音也保持高度可理解性。相反,语音中断和音高跳跃可能会干扰语音对比和正常语调,从而显著降低语音的可理解性。我们认为,NLP并没有从人类的声乐曲目中消失,而是在我们进化为语言的过程中得到了更好的控制。本文是“脊椎动物发声的非线性现象:机制和交流功能”主题的一部分。
本文章由计算机程序翻译,如有差异,请以英文原文为准。
求助全文
约1分钟内获得全文 去求助
来源期刊
CiteScore
11.80
自引率
1.60%
发文量
365
审稿时长
3 months
期刊介绍: The journal publishes topics across the life sciences. As long as the core subject lies within the biological sciences, some issues may also include content crossing into other areas such as the physical sciences, social sciences, biophysics, policy, economics etc. Issues generally sit within four broad areas (although many issues sit across these areas): Organismal, environmental and evolutionary biology Neuroscience and cognition Cellular, molecular and developmental biology Health and disease.
期刊最新文献
The effect of habitat health and environmental change on cultural diversity and richness in animals. Strategies for integrating animal social learning and culture into conservation translocation practice. Culture and conservation in baleen whales. Fishy culture in a changing world. Conserving avian vocal culture.
×
引用
GB/T 7714-2015
复制
MLA
复制
APA
复制
导出至
BibTeX EndNote RefMan NoteFirst NoteExpress
×
×
提示
您的信息不完整,为了账户安全,请先补充。
现在去补充
×
提示
您因"违规操作"
具体请查看互助需知
我知道了
×
提示
现在去查看 取消
×
提示
确定
0
微信
客服QQ
Book学术公众号 扫码关注我们
反馈
×
意见反馈
请填写您的意见或建议
请填写您的手机或邮箱
已复制链接
已复制链接
快去分享给好友吧!
我知道了
×
扫码分享
扫码分享
Book学术官方微信
Book学术文献互助
Book学术文献互助群
群 号:604180095
Book学术
文献互助 智能选刊 最新文献 互助须知 联系我们:info@booksci.cn
Book学术提供免费学术资源搜索服务,方便国内外学者检索中英文文献。致力于提供最便捷和优质的服务体验。
Copyright © 2023 Book学术 All rights reserved.
ghs 京公网安备 11010802042870号 京ICP备2023020795号-1