A new context‐dependent term weight computed by boost and discount using relevance information
E.K.F. Dang,
R.W.P. Luk,
J. Allan,
K.S. Ho,
S.C.F. Chan,
K.F.L. Chung and
D.L. Lee
Journal of the American Society for Information Science and Technology, 2010, vol. 61, issue 12, 2514-2530
Abstract:
We studied the effectiveness of a new class of context‐dependent term weights for information retrieval. Unlike the traditional term frequency–inverse document frequency (TF–IDF), the new weighting of a term t in a document d depends not only on the occurrence statistics of t alone but also on the terms found within a text window (or “document‐context”) centered on t. We introduce a Boost and Discount (B&D) procedure which utilizes partial relevance information to compute the context‐dependent term weights of query terms according to a logistic regression model. We investigate the effectiveness of the new term weights compared with the context‐independent BM25 weights in the setting of relevance feedback. We performed experiments with title queries of the TREC‐6, ‐7, ‐8, and 2005 collections, comparing the residual Mean Average Precision (MAP) measures obtained using B&D term weights and those obtained by a baseline using BM25 weights. Given either 10 or 20 relevance judgments of the top retrieved documents, using the new term weights yields improvement over the baseline for all collections tested. The MAP obtained with the new weights has relative improvement over the baseline by 3.3 to 15.2%, with statistical significance at the 95% confidence level across all four collections.
Date: 2010
References: Add references at CitEc
Citations:
Downloads: (external link)
https://doi.org/10.1002/asi.21425
Related works:
This item may be available elsewhere in EconPapers: Search for items with the same title.
Export reference: BibTeX
RIS (EndNote, ProCite, RefMan)
HTML/Text
Persistent link: https://EconPapers.repec.org/RePEc:bla:jamist:v:61:y:2010:i:12:p:2514-2530
Ordering information: This journal article can be ordered from
https://doi.org/10.1002/(ISSN)1532-2890
Access Statistics for this article
More articles in Journal of the American Society for Information Science and Technology from Association for Information Science & Technology
Bibliographic data for series maintained by Wiley Content Delivery ().