Semantic High-Level Features for Automated Cross-Modal Slideshow Generation

2009 Seventh International Workshop on Content-Based Multimedia Indexing Pub Date : 2009-06-03 DOI:10.1109/CBMI.2009.32

P. Dunker, C. Dittmar, André Begau, S. Nowak, M. Gruhne

引用次数: 8

Abstract

This paper describes a technical solution for automated slideshow generation by extracting a set of high-level features from music, such as beat grid, mood and genre and intelligently combining this set with image high-level features, such as mood, daytime- and scene classification. An advantage of this high-level concept is to enable the user to incorporate his preferences regarding the semantic aspects of music and images. For example, the user might request the system to automatically create a slideshow, which plays soft music and shows pictures with sunsets from the last 10 years of his own photo collection.The high-level feature extraction on both, the audio and the visual information is based on the same underlying machine learning core, which processes different audio- and visual- low- and mid-level features. This paper describes the technical realization and evaluation of the algorithms with suitable test databases.

查看原文

微信好友朋友圈 QQ好友复制链接

本刊更多论文

用于自动跨模态幻灯片生成的语义高级特性

本文描述了一种自动生成幻灯片的技术方案，从音乐中提取一组高级特征，如节拍网格、情绪和类型，并将其与图像高级特征(如情绪、白天和场景分类)智能结合。这种高级概念的一个优点是使用户能够结合自己对音乐和图像语义方面的偏好。例如，用户可能要求系统自动创建一个幻灯片，播放柔和的音乐，并显示最近10年他自己的照片收藏中的日落图片。本文用合适的测试数据库描述了算法的技术实现和评估。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

求助全文

约1分钟内获得全文去求助

来源期刊

2009 Seventh International Workshop on Content-Based Multimedia Indexing

自引率

0.00%

发文量