Automatic text extraction from video for content-based annotation and retrieval

27 November 2002

conference paper
Published by Institute of Electrical and Electronics Engineers (IEEE)

Vol. 1 (10514651) , 618-620
https://doi.org/10.1109/icpr.1998.711219

Abstract

Efficient content-based retrieval of image and video databases is an important application due to rapid proliferation of digital video data on the Internet and corporate intranets. Text either embedded or superimposed within video frames is very useful for describing the contents of the frames, as it enables both keyword and free-text based search, automatic video logging, and video cataloging. We have developed a scheme for automatically extracting text from digital images and videos for content annotation and retrieval. We present our approach to robust text extraction from video frames, which can handle complex image backgrounds, deal with different font sizes, font styles, and font appearances such as normal and inverse video. Our algorithm results in segmented characters that can be directly processed by an OCR system to produce ASCII text. Results from our experiments with over 5000 frames obtained from twelve MPEG video streams demonstrate the good performance of our system in terms of text identification accuracy and computational efficiency.

Keywords

This publication has 4 references indexed in Scilit:

Video skimming and characterization through the combination of image and language understanding techniques
Published by Institute of Electrical and Electronics Engineers (IEEE) ,2002
Finding text in images
Published by Association for Computing Machinery (ACM) ,1997
Automatic text recognition for video indexing
Published by Association for Computing Machinery (ACM) ,1996
Picture Thresholding Using an Iterative Selection Method
IEEE Transactions on Systems, Man, and Cybernetics, 1978