基于谱减法的计算听觉场景分析的语音增强算法

2016 IEEE International Symposium on Signal Processing and Information Technology (ISSPIT) Pub Date : 2016-12-01 DOI:10.1109/ISSPIT.2016.7886000

Cong Guo, Like Hui, Weiqiang Zhang, Jia Liu

{"title":"基于谱减法的计算听觉场景分析的语音增强算法","authors":"Cong Guo, Like Hui, Weiqiang Zhang, Jia Liu","doi":"10.1109/ISSPIT.2016.7886000","DOIUrl":null,"url":null,"abstract":"Computational auditory scene analysis (CASA) system is well used in speech enhancement area in recent years. We propose a new system that combines CASA and spectral subtraction to get better enhanced speech. The CASA part consists of the latest method deep neural networks (DNNs). The original way to reconstruct the denoise signal is to use the estimated masks with direct overlap-add method ignoring the information of noise within the frames. In our system, we estimate self-adapted thresholds for each channel by Gaussian Mixture Model from the estimated ratio masks (ERMs) to separate noise and speech of each channel. In this way, we make full use of the information within frames. The results show increase in both objective and subjective evaluation.","PeriodicalId":371691,"journal":{"name":"2016 IEEE International Symposium on Signal Processing and Information Technology (ISSPIT)","volume":"24 1","pages":"0"},"PeriodicalIF":0.0000,"publicationDate":"2016-12-01","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"3","resultStr":"{\"title\":\"A speech enhancement algorithm using computational auditory scene analysis with spectral subtraction\",\"authors\":\"Cong Guo, Like Hui, Weiqiang Zhang, Jia Liu\",\"doi\":\"10.1109/ISSPIT.2016.7886000\",\"DOIUrl\":null,\"url\":null,\"abstract\":\"Computational auditory scene analysis (CASA) system is well used in speech enhancement area in recent years. We propose a new system that combines CASA and spectral subtraction to get better enhanced speech. The CASA part consists of the latest method deep neural networks (DNNs). The original way to reconstruct the denoise signal is to use the estimated masks with direct overlap-add method ignoring the information of noise within the frames. In our system, we estimate self-adapted thresholds for each channel by Gaussian Mixture Model from the estimated ratio masks (ERMs) to separate noise and speech of each channel. In this way, we make full use of the information within frames. The results show increase in both objective and subjective evaluation.\",\"PeriodicalId\":371691,\"journal\":{\"name\":\"2016 IEEE International Symposium on Signal Processing and Information Technology (ISSPIT)\",\"volume\":\"24 1\",\"pages\":\"0\"},\"PeriodicalIF\":0.0000,\"publicationDate\":\"2016-12-01\",\"publicationTypes\":\"Journal Article\",\"fieldsOfStudy\":null,\"isOpenAccess\":false,\"openAccessPdf\":\"\",\"citationCount\":\"3\",\"resultStr\":null,\"platform\":\"Semanticscholar\",\"paperid\":null,\"PeriodicalName\":\"2016 IEEE International Symposium on Signal Processing and Information Technology (ISSPIT)\",\"FirstCategoryId\":\"1085\",\"ListUrlMain\":\"https://doi.org/10.1109/ISSPIT.2016.7886000\",\"RegionNum\":0,\"RegionCategory\":null,\"ArticlePicture\":[],\"TitleCN\":null,\"AbstractTextCN\":null,\"PMCID\":null,\"EPubDate\":\"\",\"PubModel\":\"\",\"JCR\":\"\",\"JCRName\":\"\",\"Score\":null,\"Total\":0}","platform":"Semanticscholar","paperid":null,"PeriodicalName":"2016 IEEE International Symposium on Signal Processing and Information Technology (ISSPIT)","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.1109/ISSPIT.2016.7886000","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"","JCRName":"","Score":null,"Total":0}

引用次数: 3

摘要

计算听觉场景分析(CASA)系统近年来在语音增强领域得到了很好的应用。我们提出了一种结合CASA和频谱减法的新系统，以获得更好的增强语音。CASA部分由最新方法深度神经网络(dnn)组成。原始的重建噪声信号的方法是利用直接叠加法估计的掩模，忽略帧内的噪声信息。在我们的系统中，我们使用高斯混合模型从估计的比率掩模(erm)中估计每个通道的自适应阈值，以分离每个通道的噪声和语音。这样，我们就充分利用了帧内的信息。结果表明，客观评价和主观评价均有所提高。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

查看原文

微信好友朋友圈 QQ好友复制链接

本刊更多论文

A speech enhancement algorithm using computational auditory scene analysis with spectral subtraction

Computational auditory scene analysis (CASA) system is well used in speech enhancement area in recent years. We propose a new system that combines CASA and spectral subtraction to get better enhanced speech. The CASA part consists of the latest method deep neural networks (DNNs). The original way to reconstruct the denoise signal is to use the estimated masks with direct overlap-add method ignoring the information of noise within the frames. In our system, we estimate self-adapted thresholds for each channel by Gaussian Mixture Model from the estimated ratio masks (ERMs) to separate noise and speech of each channel. In this way, we make full use of the information within frames. The results show increase in both objective and subjective evaluation.

求助全文

通过发布文献求助，成功后即可免费获取论文全文。去求助

来源期刊

2016 IEEE International Symposium on Signal Processing and Information Technology (ISSPIT)

自引率

0.00%

发文量

期刊最新文献

Informed Split Gradient Non-negative Matrix factorization using Huber cost function for source apportionment An Identity and Access Management approach for SOA Extracting dispersion information from Optical Coherence Tomography images LOS millimeter-wave communication with quadrature spatial modulation An FPGA design for the Two-Band Fast Discrete Hartley Transform