您当前所在位置: 首页 > 学者

周志华

  • 36浏览

  • 0点赞

  • 0收藏

  • 0分享

  • 73下载

  • 0评论

  • 引用

期刊论文

Multi-Instance Learning Based Web Mining

周志华ZHI-HUA ZHOU* KAI JIANG AND MING LI

Applied Intelligence 22, 135-147, 2005,-0001,():

URL:

摘要/描述

In multi-instance learning, the training set comprises labeled bags that are composed of unlabeled instances, and the task is to predict the labels of unseen bags. In this paper, a web mining problem, i.e. web index recommendation, is investigated from a multi-instance view. In detail, each web index page is regarded as a bag, while each of its linked pages is regarded as an instance. A user favoring an index page means that he or she is interested in at least one page linked by the index. Based on the browsing history of the user, recommendation could be provided for unseen index pages. An algorithm named Fretcit-kNN, which employs the Minimal Hausdorff distance between frequent term sets and utilizes both the references and citers of an unseen bag in determining its label, is proposed to solve the problem. Experiments show that in average the recommendation accuracy of Fretcit-kNN is 81.0% with 71.7% recall and 70.9% precision, which is significantly better than the best algorithm that does not consider the specific characteristics of multi-instance learning, whose performance is 76.3% accuracy with 63.4% recall and 66.1% precision.

版权说明:以下全部内容由周志华上传于   2005年08月02日 17时44分56秒,版权归本人所有。

我要评论

全部评论 0

本学者其他成果

    同领域成果