A novel video summarization based on mining the story-structure and semantic relations among concept entities

Bo Wei Chen, Jia Ching Wang, Jhing Fa Wang

研究成果: 雜誌貢獻期刊論文同行評審

115 引文 斯高帕斯(Scopus)

摘要

Video summarization techniques have been proposed for years to offer people comprehensive understanding of the whole story in the video. Roughly speaking, existing approaches can be classified into the two types: one is static storyboard, and the other is dynamic skimming. However, despite that these traditional methods give brief summaries for users, they still do not provide with a concept-organized and systematic view. In this paper, we present a structural video content browsing system and a novel summarization method by utilizing the four kinds of entities: who, what, where, and when to establish the framework of the video contents. With the assistance of the above-mentioned indexed information, the structure of the story can be built up according to the characters, the things, the places, and the time. Therefore, users can not only browse the video efficiently but also focus on what they are interested in via the browsing interface. In order to construct the fundamental system, we employ maximum entropy criterion to integrate visual and text features extracted from video frames and speech transcripts, generating high-level concept entities. A novel concept expansion method is introduced to explore the associations among these entities. After constructing the relational graph, we exploit graph entropy model to detect meaningful shots and relations, which serve as the indices for users. The results demonstrate that our system can achieve better performance and information coverage.

原文???core.languages.en_GB???
文章編號4757424
頁(從 - 到)295-312
頁數18
期刊IEEE Transactions on Multimedia
11
發行號2
DOIs
出版狀態已出版 - 2月 2009

指紋

深入研究「A novel video summarization based on mining the story-structure and semantic relations among concept entities」主題。共同形成了獨特的指紋。

引用此