Multiple-description coding (MDC) of speech with an invertible auditory model

1999 IEEE Workshop on Speech Coding Proceedings. Model, Coders, and Error Criteria (Cat. No.99EX351) Pub Date : 1999-06-20 DOI:10.1109/SCFT.1999.781491

G. Kubin, W. Kleijn

引用次数: 18

Abstract

Network signal processing aspects dominate in speech and audio coding applications such as Internet telephony or packet radio networks. We demonstrate that our approach to speech coding in a perceptual domain provides an implicit forward error concealment mechanism to handle random erasures of the channel. To this end, the individual acoustic subchannels of our auditory model are grouped into different transport subchannels or packets. Due to the strongly overlapping, redundant filterbank structure of the model, reconstruction of speech without audible degradation becomes possible even if a significant percentage of channels is erased (e.g., up to 40% in a 50-channel auditory model for narrowband speech). We discuss this result both from a hearing-physiology and a frame-theoretic perspective.

查看原文

微信好友朋友圈 QQ好友复制链接

本刊更多论文

基于可逆听觉模型的语音多描述编码

网络信号处理方面在语音和音频编码应用中占主导地位，如互联网电话或分组无线网络。我们证明了我们在感知域的语音编码方法提供了一种隐式的前向错误隐藏机制来处理信道的随机擦除。为此，我们的听觉模型的单个声学子通道被分组到不同的传输子通道或数据包中。由于模型的强重叠冗余滤波器组结构，即使有很大比例的通道被擦除(例如，在50通道的窄带语音听觉模型中高达40%)，也可以在没有听觉退化的情况下重建语音。我们从听觉生理学和框架理论的角度来讨论这一结果。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

求助全文

约1分钟内获得全文去求助

来源期刊

1999 IEEE Workshop on Speech Coding Proceedings. Model, Coders, and Error Criteria (Cat. No.99EX351)

自引率

0.00%

发文量

期刊最新文献

Reverse water-filling in predictive encoding of speech Integration of speech enhancement and coding techniques The use of LSF-based phonetic classification in low-rate coder design Post noise smoother to improve low bit rate speech-coding performance A novel pitch-lag search method using adaptive weighting and median filtering