首页 | 本学科首页   官方微博 | 高级检索  
     


Document Clustering With Dual Supervision Through Feature Reweighting
Authors:Yeming Hu  Evangelos E Milios  James Blustein
Affiliation:1. Faculty of Computer Science, Dalhousie University, Nova Scotia, Canada;2. Faculty of Computer Science and School of Management, Dalhousie University, Nova Scotia, Canada
Abstract:Traditional semi‐supervised clustering uses only limited user supervision in the form of instance seeds for clusters and pairwise instance constraints to aid unsupervised clustering. However, user supervision can also be provided in alternative forms for document clustering, such as labeling a feature by indicating whether it discriminates among clusters. This article thus fills this void by enhancing traditional semi‐supervised clustering with feature supervision, which asks the user to label discriminating features during defining (labeling) the instance seeds or pairwise instance constraints. Various types of semi‐supervised clustering algorithms were explored with feature supervision. Our experimental results on several real‐world data sets demonstrate that augmenting the instance‐level supervision with feature‐level supervision can significantly improve document clustering performance.
Keywords:user supervision  feature supervision  feature reweighting  text cloud
设为首页 | 免责声明 | 关于勤云 | 加入收藏

Copyright©北京勤云科技发展有限公司  京ICP备09084417号