DocumentCode
2264900
Title
VSUM: summarizing from videos
Author
Wu, Yu-Chyeh ; Lee, Yue-Shi ; Chang, Chia-Hui
Author_Institution
Nat. Central Univ., Taoyuan, Taiwan
fYear
2004
fDate
13-15 Dec. 2004
Firstpage
302
Lastpage
309
Abstract
Summarization on produced type of video data (like news or movies) is to find important segments that contain rich information. Users could obtain the important messages by reading summaries rather than full documents. The research in this area could be divided into two parts: (I) image processing (IP) perspective, and (2) NLP (nature language processing) perspective. The former put emphasis on the detection of key frames, while the later focused on the extraction of important concepts. This paper proposes a video summarization system, VSUM. VSUM first identifies all caption words, and then adopts a technique to find the important segments. An external thesaurus is also used in VSUM to enhance the summary extraction process. The experimental results show that VSUM could perform well even if the accuracy of OCR (optical character recognition) is not sophisticating.
Keywords
image processing; natural languages; optical character recognition; video databases; VSUM; image processing perspective; movies; nature language processing perspective; news; optical character recognition; video data; video summarization system; Biomedical imaging; Cities and towns; Image processing; Image segmentation; Motion detection; Motion pictures; Optical character recognition software; Pattern recognition; Thesauri; Videos;
fLanguage
English
Publisher
ieee
Conference_Titel
Multimedia Software Engineering, 2004. Proceedings. IEEE Sixth International Symposium on
Print_ISBN
0-7695-2217-3
Type
conf
DOI
10.1109/MMSE.2004.90
Filename
1376676
Link To Document