Automatic story segmentation for spoken document retrieval
Refereed conference paper presented and published in conference proceedings


全文

引用次數

其它資訊
摘要We have been working on speech retrieval based on Cantonese television news programs. Our video archive contains over 20 hours of news programs provided by a local television station. These programs have been hand-segmented into video clips, where each clip is a self-contained news story. The audio tracks in our archive are indexed by Cantonese speech recognition. This is integrated with a vector-space information retrieval model to achieve speech retrieval. This paper proposes an approach for automatic story segmentation from television news programs, intended to replace hand-segmentation as described above. Automatic story segmentation is critical for rapid expansion of out video archive. Our approach relies on the assumption that nearly all the news stories follow the temporal syntax of (begin_story --> anchor shots --> field shots --> end_story). Therefore our algorithm aims to detect field-to-anchor shot boundaries, that should also coincide with the story boundaries. The proposed approach utilizes the video frame information for story boundary detection, and involves such techniques as fuzzy c-means and graph-theoretical clustering. The approach achieved precision and recall values of over 70%, based on a 20-hour video corpus.
著者Hui PY, Tang XO, Meng HM, Lam W, Gao XB
會議名稱10th IEEE International Conference on Fuzzy Systems
會議開始日02.12.2001
會議完結日05.12.2001
會議地點MELBOURNE
會議國家/地區澳大利亞
出版年份2001
月份1
日期1
出版社IEEE
頁次1319 - 1322
國際標準書號0-7803-7293-X
語言英式英語
Web of Science 學科類別Automation & Control Systems; Computer Science; Computer Science, Artificial Intelligence; Engineering; Engineering, Electrical & Electronic; Imaging Science & Photographic Technology

上次更新時間 2020-20-09 於 03:23