Chapter

Information Retrieval Technology

Volume 4182 of the series Lecture Notes in Computer Science pp 132-144

Text Clustering with Limited User Feedback Under Local Metric Learning

  • Ruizhang HuangAffiliated withCarnegie Mellon UniversityDepartment of Systems Engineering and Engineering Management, The Chinese University of Hong Kong
  • , Zhigang ZhangAffiliated withCarnegie Mellon UniversityDepartment of Systems Engineering and Engineering Management, The Chinese University of Hong Kong
  • , Wai LamAffiliated withCarnegie Mellon UniversityDepartment of Systems Engineering and Engineering Management, The Chinese University of Hong Kong

* Final gross prices may vary according to local VAT.

Get Access

Abstract

This paper investigates the idea of incorporating incremental user feedbacks and a small amount of sample documents for some, not necessarily all, clusters into text clustering. For the modeling of each cluster, we make use of a local weight metric to reflect the importance of the features for a particular cluster. The local weight metric is learned using both the unlabeled data and the constraints generated automatically from user feedbacks and sample documents. The quality of local metric is improved by incorporating more precise constraints. Improving the quality of local metric will in return enhance the clustering performance. We have conducted extensive experiments on real-world news documents. The results demonstrate that user feedback information coupled with local metric learning can dramatically improve the clustering performance.