Sciweavers

10500 search results - page 282 / 2100
» Documentation for
Sort
View
177
Voted
ECIR
2009
Springer
16 years 4 months ago
Evaluation of Text Clustering Algorithms with N-Gram-Based Document Fingerprints
This paper presents a new approach designed to reduce the computational load of the existing clustering algorithms by trimming down the documents size using fingerprinting methods...
Javier Parapar, Alvaro Barreiro
191
Voted
CIKM
2009
Springer
16 years 2 months ago
The impact of document structure on keyphrase extraction
Keyphrases are short phrases that reflect the main topic of a document. Because manually annotating documents with keyphrases is a time-consuming process, several automatic appro...
Katja Hofmann, Manos Tsagkias, Edgar Meij, Maarten...
186
Voted
ICDM
2007
IEEE
143views Data Mining» more  ICDM 2007»
16 years 1 months ago
Bit Sequences and Biclustering of Text Documents
We propose a new technique for clustering of text documents that relies on a biclustering structure constructed on terms and documents. Our approach makes use of a greedy algorith...
Selim Mimaroglu, Kuniaki Uehara
ELPUB
2006
ACM
16 years 1 months ago
A View on Two Complementary Representations of Documents for Information Retrieval
The indexation of documents is a critical step of the information retrieval process and is often a manual task which highly depends on the indexer’s knowledge. We propose to imp...
Béatrice Rumpler, Hassan Naderi
179
Voted
JCDL
2005
ACM
100views Education» more  JCDL 2005»
16 years 1 months ago
Automatic extraction of titles from general documents using machine learning
In this paper, we propose a machine learning approach to title extraction from general documents. By general documents, we mean documents that can belong to any one of a number of...
Yunhua Hu, Hang Li, Yunbo Cao, Dmitriy Meyerzon, Q...