Sciweavers

5046 search results - page 684 / 1010
» Non-redundant data clustering
Sort
View
TKDE
2012
190views Formal Methods» more  TKDE 2012»
13 years 10 months ago
Scalable Learning of Collective Behavior
—This study of collective behavior is to understand how individuals behave in a social networking environment. Oceans of data generated by social media like Facebook, Twitter, Fl...
Lei Tang, Xufei Wang, Huan Liu
ICSE
2012
IEEE-ACM
13 years 10 months ago
What make long term contributors: Willingness and opportunity in OSS community
—To survive and succeed, software projects need to attract and retain contributors. We model the individual’s chances to become a valuable contributor through her capacity, wil...
Minghui Zhou, Audris Mockus
ICSE
2012
IEEE-ACM
13 years 10 months ago
Synthesizing API usage examples
Abstract—Key program interfaces are sometimes documented with usage examples: concrete code snippets that characterize common use cases for a particular data type. While such doc...
Raymond P. L. Buse, Westley Weimer
KDD
2008
ACM
156views Data Mining» more  KDD 2008»
16 years 8 months ago
Unsupervised deduplication using cross-field dependencies
Recent work in deduplication has shown that collective deduplication of different attribute types can improve performance. But although these techniques cluster the attributes col...
Robert Hall, Charles A. Sutton, Andrew McCallum
KDD
2004
ACM
158views Data Mining» more  KDD 2004»
16 years 8 months ago
A generalized maximum entropy approach to bregman co-clustering and matrix approximation
Co-clustering is a powerful data mining technique with varied applications such as text clustering, microarray analysis and recommender systems. Recently, an informationtheoretic ...
Arindam Banerjee, Inderjit S. Dhillon, Joydeep Gho...