OCR Based Slide Retrieval

Daddaoua, N., Odobez, J.M. and Vinciarelli, A. (2005) OCR Based Slide Retrieval. In: ICDAR 2005 Proceedings of the Eighth International Conference on Document Analysis and Recognition, Seoul, South Korea, 29 August - 1 September, pp. 945-949. (doi: 10.1109/ICDAR.2005.169)

Full text not currently available from Enlighten.

Abstract

This paper addresses the problem of acquiring, indexing and retrieving slides in the context of automatic oral presentation processing. Since the most suitable acquisition technique, in such a context, is the use of a framegrabber (a device capturing as images the slides displayed on a screen), the slides must be transcribed with an optical character recognition system. Retrieval experiments performed on a corpus of 570 slides (26 presentations) gathered at a workshop show that performance obtained with the OCR transcriptions are close to those obtained by extracting the text from the electronic version (pdf or ppt) of the slides (through apposite APIs).

Item Type:Conference Proceedings
Status:Published
Refereed:Yes
Glasgow Author(s) Enlighten ID:Vinciarelli, Professor Alessandro
Authors: Daddaoua, N., Odobez, J.M., and Vinciarelli, A.
College/School:College of Science and Engineering > School of Computing Science

University Staff: Request a correction | Enlighten Editors: Update this record