Skip to main navigation Skip to search Skip to main content

Novelty detection for cross-lingual news stories with visual duplicates and speech transcripts

Research output: Chapters, Conference Papers, Creative and Literary WorksRGC 32 - Refereed conference paper (with host publication)peer-review

Abstract

An overwhelming volume of news videos from different channels and languages is available today, which demands automatic management of this abundant information. To effectively search, retrieve, browse and track cross-lingual news stories, a news story similarity measure plays a critical role in assessing the novelty and redundancy among them. In this paper, we explore the novelty and redundancy detection with visual duplicates and speech transcripts for cross-lingual news stories. News stories are represented by a sequence of keyframes in the visual track and a set of words extracted from speech transcript in the audio track. A major difference to pure text documents is that the number of keyframes in one story is relatively small compared to the number of words and there exist a large number of non-near duplicate keyframes. These features make the behavior of similarity measures different compared to traditional textual collections. Furthermore, the textual features and visual features complement each other for news stories. They can be further combined to boost the performance. Experiments on the TRECVID-2005 cross-lingual news video corpus show that approaches on textual features and visual features demonstrate different performance, and measures on visual features are quite effective. Overall, the cosine distance on keyframes is still a robust measure. Language models built on visual features demonstrate promising performance. The fusion of textual and visual features improves overall performance. Copyright 2007 ACM.
Original languageEnglish
Title of host publicationProceedings of the ACM International Multimedia Conference and Exhibition
Pages168-177
DOIs
Publication statusPublished - 2007
Event15th ACM International Conference on Multimedia, MM'07 - Augsburg, Bavaria, Germany
Duration: 24 Sept 200729 Sept 2007

Conference

Conference15th ACM International Conference on Multimedia, MM'07
PlaceGermany
CityAugsburg, Bavaria
Period24/09/0729/09/07

Research Keywords

  • Cross-lingual information retrieval
  • Language model
  • Multimodality
  • Near-duplicate keyframes
  • News videos
  • Novelty and redundancy detection
  • Similarity measure

Fingerprint

Dive into the research topics of 'Novelty detection for cross-lingual news stories with visual duplicates and speech transcripts'. Together they form a unique fingerprint.

Cite this