Speech-to-text input method for web system using JavaScript

2008 IEEE Spoken Language Technology Workshop Pub Date : 2008-12-01 DOI:10.1109/SLT.2008.4777877

R. Nisimura, Jumpei Miyake, Hideki Kawahara, T. Irino

引用次数: 14

Abstract

We have developed a speech-to-text input method for web systems. The system is provided as a JavaScript library including an Ajax-like mechanism based on a Java applet, CGI programs, and dynamic HTML documents. It allows users to access voice-enabled web pages without requiring special browsers. Web developers can embed it on their web page by inserting only one line in the header field of an HTML document. This study also aims at observing natural spoken interactions in personal environments. We have succeeded in collecting 4,003 inputs during a period of seven months via our public Japanese ASR server. In order to cover out-of-vocabulary words to cope with some proper nouns, a web page to register new words into the language model are developed. As a result, we could obtain an improvement of 0.8% in the recognition accuracy. With regard to the acoustical conditions, an SNR of 25.3 dB was observed.

查看原文

微信好友朋友圈 QQ好友复制链接

本刊更多论文

语音转文本输入法的web系统，使用JavaScript

我们为网络系统开发了一种语音到文本的输入法。该系统以JavaScript库的形式提供，包括基于Java applet的类ajax机制、CGI程序和动态HTML文档。它允许用户在不需要特殊浏览器的情况下访问语音网页。Web开发人员只需在HTML文档的标题字段中插入一行，就可以将其嵌入到网页中。本研究还旨在观察个人环境中的自然言语互动。在7个月的时间里，我们通过日本公共ASR服务器成功收集了4003个输入。为了覆盖词汇外的单词，以应对一些专有名词，开发了一个网页，将新单词注册到语言模型中。结果表明，我们的识别准确率提高了0.8%。声学条件下，信噪比为25.3 dB。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

求助全文

约1分钟内获得全文去求助

来源期刊

2008 IEEE Spoken Language Technology Workshop

自引率

0.00%

发文量

期刊最新文献

“Who is this” quiz dialogue system and users' evaluation Latent dirichlet language model for speech recognition Modelling user behaviour in the HIS-POMDP dialogue manager A syntactic language model based on incremental CCG parsing Improving word segmentation for Thai speech translation