A similar content retrieval method for podcast episodes

2008 IEEE Spoken Language Technology Workshop Pub Date : 2008-12-01 DOI:10.1109/SLT.2008.4777899

Junta Mizuno, J. Ogata, Masataka Goto

引用次数: 10

Abstract

Given podcasts (audio blogs) that are sets of speech files called episodes, this paper describes a method for retrieving episodes that have similar content. Although most previous retrieval methods were based on bibliographic information, tags, or users' playback behaviors without considering spoken content, our method can compute content-based similarity based on speech recognition results of podcast episodes even if the recognition results include some errors. To overcome those errors, it converts intermediate speech-recognition results to a confusion network containing competitive candidates, and then computes the similarity by using keywords extracted from the network. Experimental results with episodes that have different word accuracy and content showed that keywords obtained from competitive candidates were useful in retrieving similar episodes. To show relevant episodes, our method will be incorporated into PodCastle, a public web service that provides full-text searching of podcasts on the basis of speech recognition.

查看原文

微信好友朋友圈 QQ好友复制链接

本刊更多论文

播客集的类似内容检索方法

鉴于播客(音频博客)是一组称为剧集的语音文件，本文描述了一种检索具有相似内容的剧集的方法。虽然以前的大多数检索方法是基于书目信息、标签或用户的播放行为而不考虑语音内容，但我们的方法可以基于播客集的语音识别结果计算基于内容的相似度，即使识别结果存在一些错误。为了克服这些错误，它将中间语音识别结果转换为包含竞争候选人的混淆网络，然后使用从网络中提取的关键字计算相似度。实验结果表明，从竞争候选词中获得的关键词在检索相似的剧集时是有用的。为了显示相关的剧集，我们的方法将被整合到PodCastle中。PodCastle是一个公共网络服务，提供基于语音识别的播客全文搜索。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

求助全文

约1分钟内获得全文去求助

来源期刊

2008 IEEE Spoken Language Technology Workshop

自引率

0.00%

发文量

期刊最新文献

“Who is this” quiz dialogue system and users' evaluation Latent dirichlet language model for speech recognition Modelling user behaviour in the HIS-POMDP dialogue manager A syntactic language model based on incremental CCG parsing Improving word segmentation for Thai speech translation