A Novel Video Caption Detection Approach Using Multi-Frame Integration

RR Wang,WJ Jin,LD Wu
DOI: https://doi.org/10.1109/icpr.2004.1334156
2005-01-01
Journal of Computer Research and Development
Abstract:Captions in videos often play an important role in video information indexing and retrieval. In this paper, we present a novel video caption detection approach. We first apply a new multiple frame integration (MFI) method to minimize variation of the background of the image. A time-based minimum (or maximum) pixel value search is employed and a Sobel edge map is used to determine the mode of search. Then block-based text detection is performed, i.e., a small window is used to scan the image and classify as text or non-text, using Sobel edges as features. We use a two-level pyramid to detect various text sizes. Finally, we present a new iterative text line decomposition method and accurate text bounding boxes are extracted from candidate text areas. Experimental result shows that the proposed approach achieves high precision and recall.
What problem does this paper attempt to address?