Sciweavers

10500 search results - page 361 / 2100
» Documentation for
Sort
View
ICDAR
2009
IEEE
16 years 2 months ago
Automated Ground Truth Data Generation for Newspaper Document Images
In document image understanding, public datasets with ground-truth are an important part of scientific work. They are not only helpful for developing new methods, but also provid...
Thomas Strecker, Joost van Beusekom, Sahin Albayra...
FQAS
2009
Springer
129views Database» more  FQAS 2009»
16 years 2 months ago
Representing Context Information for Document Retrieval
The bag of words representation (BoW), which is widely used in information retrieval (IR), represents documents and queries as word lists that do not express anything about context...
Maya Carrillo, Esaú Villatoro-Tello, Aureli...
ICPR
2008
IEEE
16 years 2 months ago
Ancient document analysis based on text line extraction
In order to preserve our cultural heritage and for automated document processing libraries and national archives have started digitizing historical documents. In the case of degra...
Florian Kleber, Robert Sablatnig, Melanie Gau, Hei...
ICPR
2008
IEEE
16 years 2 months ago
A robust technique for text extraction in mixed-type binary documents
A crucial preprocessing stage in applications such as OCR is text extraction from mixed-type documents. The present work, in contrast to most until now, successfully faces the pro...
Charalambos Strouthopoulos, Athanasios Nikolaidis
ICDAR
2007
IEEE
16 years 1 months ago
A Sparse and Locally Shift Invariant Feature Extractor Applied to Document Images
We describe an unsupervised learning algorithm for extracting sparse and locally shift-invariant features. We also devise a principled procedure for learning hierarchies of invari...
Marc'Aurelio Ranzato, Yann LeCun