Multi-modal Interview Concept Detection for Rushes Exploitation

An-An Liu, Sheng Tang, Yongdong Zhang, Jintao Li, Zhaoxuan Yang · 2009

According to the concepts of Large-Scale Concept Ontology for Multimedia (LSCOM) and requirement of the 4th task in the 2006 TRECVID, i.e., rushes exploitation, the “interview ” concept is an important semantic concept for rushes content analysis. The paper presents the shot-level “interview ” concept detection method. Face detection and audio classification are implemented to detect “face ” and “speech ” concepts for each shot. By integrating audiovisual information, “interview ” concept is finally detected. The utilization of the method will definitely benefit the video edit. Large-scale experimental results strongly demonstrate the accuracy and effectiveness of the proposed method. The TREC conference series is sponsored by the National Institute of Standards and Technology (NIST) with additional support from other U.S. government agencies. The goal of the conference series is to encourage research in information retrieval (Guidelines, 2006). In the 2006 TRECVID, there are three system tasks and one exploratory task: shot boundary

Read the paper · More papers on PaperTik