2003 IEEE Computer Society Conference on Computer Vision and Pattern Recognition (CVPR '03) - Volume 2
Video Content Annotation Using Visual Analysis and a Large Semantic Knowledgebase
Madison, Wisconsin
June 18-June 20
ISBN: 0-7695-1900-8
We present a novel approach to automatically annotating broadcast video. To manage the enormous variety of objects, events and scenes in video problem domains such as news video, we couple generic image analysis with a semantic database, WordNet, containing huge amounts of real-world information. Object and event recognition are performed by searching WordNet for concepts jointly supported by image evidence and topic context derived from the video transcript. No object- or event-specific training is required, and only a few object models and detection algorithms are required to label much of the significant content of news video. The hierarchical structure of WordNet yields hierarchical recognition, dynamically tailored to the level of supporting image evidence. The potential of the approach is demonstrated by analyzing a wide variety of scenes in news video.
Citation:
Anthony Hoogs, Jens Rittscher, Gees Stein, John Schmiederer, "Video Content Annotation Using Visual Analysis and a Large Semantic Knowledgebase," cvpr, vol. 2, pp.327, 2003 IEEE Computer Society Conference on Computer Vision and Pattern Recognition (CVPR '03) - Volume 2, 2003