Sciweavers

6545 search results - page 916 / 1309
» Data quality assessment
Sort
View
ICDM
2009
IEEE
153views Data Mining» more  ICDM 2009»
15 years 5 months ago
A New Clustering Algorithm Based on Regions of Influence with Self-Detection of the Best Number of Clusters
Clustering methods usually require to know the best number of clusters, or another parameter, e.g. a threshold, which is not ever easy to provide. This paper proposes a new graph-b...
Fabrice Muhlenbach, Stéphane Lallich
BIB
2010
103views more  BIB 2010»
15 years 5 months ago
Challenges of sequencing human genomes
Massively parallel sequencing technologies continue to alter the study of human genetics. As the cost of sequencing declines, next-generation sequencing (NGS) instruments and data...
Daniel C. Koboldt, Li Ding, Elaine R. Mardis, Rich...
ADC
2010
Springer
204views Database» more  ADC 2010»
15 years 2 months ago
Systematic clustering method for l-diversity model
Nowadays privacy becomes a major concern and many research efforts have been dedicated to the development of privacy protecting technology. Anonymization techniques provide an eff...
Md. Enamul Kabir, Hua Wang, Elisa Bertino, Yunxian...
CORR
2011
Springer
177views Education» more  CORR 2011»
15 years 2 months ago
Tuffy: Scaling up Statistical Inference in Markov Logic Networks using an RDBMS
Markov Logic Networks (MLNs) have emerged as a powerful framework that combines statistical and logical reasoning; they have been applied to many data intensive problems including...
Feng Niu, Christopher Ré, AnHai Doan, Jude ...
NAR
2011
274views Computer Vision» more  NAR 2011»
15 years 2 months ago
A series of PDB related databases for everyday needs
The Protein Data Bank (PDB) is the world-wide repository of macromolecular structure information. We present a series of databases that run parallel to the PDB. Each database hold...
Robbie P. Joosten, Tim A. H. te Beek, Elmar Kriege...