Skip to main navigation Skip to search Skip to main content

HIT-LTRC at TREC 2010 blog track: Faceted blog distillation

  • Jinfeng Yang*
  • , Xishuang Dong
  • , Yi Guan
  • , Chengzhen Huang
  • , Sheng Wang
  • *Corresponding author for this work
  • School of Computer Science and Technology, Harbin Institute of Technology

Research output: Contribution to journalConference articlepeer-review

Abstract

This paper describes our participation in the faceted blog distillation task at Blog Track 2010. In our approach, indri toolkit is applied for basic topic relevance retrieval. Then the Maximum Entropy (ME) model is adopted to judge the relevance of each blog to specified facet. Feed faceted relevance is calculated by integrating the average relevance of all blogs within a feed and the average relevance of the most relevant N blogs. Two implementations are applied to calculate feed faceted relevance. Experiental results on Blogs08 dataset show the effectiveness of our approach.

Original languageEnglish
JournalNIST Special Publication
StatePublished - 2010
Externally publishedYes
Event19th Text REtrieval Conference, TREC 2010 - Gaithersburg, MD, United States
Duration: 16 Nov 201019 Nov 2010

Cite this