Automatic text extraction from video for content-based annotation and retrieval
- 27 November 2002
- conference paper
- Published by Institute of Electrical and Electronics Engineers (IEEE)
- Vol. 1 (10514651) , 618-620
- https://doi.org/10.1109/icpr.1998.711219
Abstract
Efficient content-based retrieval of image and video databases is an important application due to rapid proliferation of digital video data on the Internet and corporate intranets. Text either embedded or superimposed within video frames is very useful for describing the contents of the frames, as it enables both keyword and free-text based search, automatic video logging, and video cataloging. We have developed a scheme for automatically extracting text from digital images and videos for content annotation and retrieval. We present our approach to robust text extraction from video frames, which can handle complex image backgrounds, deal with different font sizes, font styles, and font appearances such as normal and inverse video. Our algorithm results in segmented characters that can be directly processed by an OCR system to produce ASCII text. Results from our experiments with over 5000 frames obtained from twelve MPEG video streams demonstrate the good performance of our system in terms of text identification accuracy and computational efficiency.Keywords
This publication has 4 references indexed in Scilit:
- Video skimming and characterization through the combination of image and language understanding techniquesPublished by Institute of Electrical and Electronics Engineers (IEEE) ,2002
- Finding text in imagesPublished by Association for Computing Machinery (ACM) ,1997
- Automatic text recognition for video indexingPublished by Association for Computing Machinery (ACM) ,1996
- Picture Thresholding Using an Iterative Selection MethodIEEE Transactions on Systems, Man, and Cybernetics, 1978