Bibliographie complète
Using query logs to establish vocabularies in distributed information retrieval
Type de ressource
Auteurs/contributeurs
- Shokouhi, Milad (Auteur)
- Zobel, Justin (Auteur)
- Tahaghoghi, Saied (Auteur)
- Scholer, Falk (Auteur)
Titre
Using query logs to establish vocabularies in distributed information retrieval
Résumé
Users of search engines express their needs as queries, typically consisting of a small number of terms. The resulting search engine query logs are valuable resources that can be used to predict how people interact with the search system. In this paper, we introduce two novel applications of query logs, in the context of distributed information retrieval. First, we use query log terms to guide sampling from uncooperative distributed collections. We show that while our sampling strategy is at least as efficient as current methods, it consistently performs better. Second, we propose and evaluate a pruning strategy that uses query log information to eliminate terms. Our experiments show that our proposed pruning method maintains the accuracy achieved by complete indexes, while decreasing the index size by up to 60%. While such pruning may not always be desirable in practice, it provides a useful benchmark against which other pruning strategies can be measured.
Publication
Information Processing & Management
Volume
43
Numéro
1
Pages
169-180
Date
janvier 2007
Abrév. de revue
Information Processing & Management
ISSN
0306-4573
Titre abrégé
Using query logs to establish vocabularies in distributed information retrieval
Consulté le
2016-07-27 17 h 23
Catalogue de bibl.
ScienceDirect
Référence
Shokouhi, M., Zobel, J., Tahaghoghi, S. et Scholer, F. (2007). Using query logs to establish vocabularies in distributed information retrieval. Information Processing & Management, 43(1), 169‑180. https://doi.org/10.1016/j.ipm.2006.04.003
Recherches connexes
Lien vers cette notice