WikiSpeech - enabling open source text-to-speech for Wikipedia

Speech Synthesis Workshop Pub Date : 2016-09-13 DOI:10.21437/SSW.2016-16

J. Andersson, S. Berlin, André Costa, Harald Berthelsen, Hanna Lindgren, N. Lindberg, J. Beskow, Jens Edlund, Joakim Gustafson

引用次数: 4

Abstract

We present WikiSpeech, an ambitious joint project aiming to (1) make open source text-to-speech available through Wikimedia Foundation’s server architecture; (2) utilize the large and active Wikipedia user base to achieve continuously improving text-to-speech; (3) improve existing and develop new crowdsourcing methods for text-to-speech; and (4) develop new and adapt current evaluation methods so that they are well suited for the particular use case of reading Wikipedia articles out loud while at the same time capable of harnessing the huge user base made available by Wikipedia. At its inauguration, the project is backed by The Swedish Post and Telecom Authority and headed by Wikimedia Sverige, STTS and KTH, but in the long run, the project aims at broad multinational involvement. The vision of the project is freely available text-to-speech for all Wikipedia languages (currently 293). In this paper, we present the project itself and its first steps: requirements, initial architecture, and initial steps to include crowdsourcing and evaluation.

查看原文

微信好友朋友圈 QQ好友复制链接

本刊更多论文

WikiSpeech -为维基百科启用开源文本到语音

我们提出WikiSpeech，这是一个雄心勃勃的联合项目，旨在(1)通过维基媒体基金会的服务器架构提供开源的文本到语音;(2)利用维基百科庞大而活跃的用户群，实现文本到语音的不断改进;(3)完善现有的文本转语音众包方式，并开发新的众包方式;(4)开发新的和调整现有的评估方法，使它们非常适合大声阅读维基百科文章的特定用例，同时能够利用维基百科提供的庞大用户群。在启动时，该项目得到了瑞典邮政和电信管理局的支持，并由维基媒体公司、STTS和KTH领导，但从长远来看，该项目旨在广泛的跨国参与。该项目的愿景是免费提供所有维基百科语言的文本到语音(目前有293种)。在本文中，我们介绍了项目本身及其第一步:需求，初始架构，以及包括众包和评估的初始步骤。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

求助全文

约1分钟内获得全文去求助

来源期刊

Speech Synthesis Workshop

自引率

0.00%

发文量