Sciweavers

6388 search results - page 412 / 1278
» High Performance Data Mining
Sort
View
KDD
2006
ACM
166views Data Mining» more  KDD 2006»
16 years 8 months ago
Anonymizing sequential releases
An organization makes a new release as new information become available, releases a tailored view for each data request, releases sensitive information and identifying information...
Ke Wang, Benjamin C. M. Fung
KDD
2010
ACM
218views Data Mining» more  KDD 2010»
15 years 11 months ago
Online multiscale dynamic topic models
We propose an online topic model for sequentially analyzing the time evolution of topics in document collections. Topics naturally evolve with multiple timescales. For example, so...
Tomoharu Iwata, Takeshi Yamada, Yasushi Sakurai, N...
SDM
2007
SIAM
137views Data Mining» more  SDM 2007»
15 years 9 months ago
Are approximation algorithms for consensus clustering worthwhile?
Consensus clustering has emerged as one of the principal clustering problems in the data mining community. In recent years the theoretical computer science community has generated...
Michael Bertolacci, Anthony Wirth
SDM
2007
SIAM
152views Data Mining» more  SDM 2007»
15 years 9 months ago
HP2PC: Scalable Hierarchically-Distributed Peer-to-Peer Clustering
In distributed data mining models, adopting a flat node distribution model can affect scalability. To address the problem of modularity, flexibility and scalability, we propose...
Khaled M. Hammouda, Mohamed S. Kamel
BMCBI
2007
128views more  BMCBI 2007»
15 years 7 months ago
Detailed estimation of bioinformatics prediction reliability through the Fragmented Prediction Performance Plots
Background: An important and yet rather neglected question related to bioinformatics predictions is the estimation of the amount of data that is needed to allow reliable predictio...
Oliviero Carugo