Español English Contacte con nosotros http://www.uc3m.es/portal/page/portal/biblioteca
DSpace e-Archivo

Archivo Abierto Institucional de la Universidad Carlos III de Madrid > Investigación > Departamentos > Departamento de Informática > Grupo de Investigación en Planificación y Aprendizaje Automático (PLG) > DI - PLG - Artículos de Revistas >

Please use this identifier to cite or link to this item: http://hdl.handle.net/10016/6744

Google™ Scholar. Others By: Billhardt, Holger - Borrajo, Daniel - Maojo, Víctor
Files in This Item:
billhardt_learning_IJU_2003_ps.pdf1,09 MBAdobe PDFformato pdf
Title: Learning retrieval expert combinations with genetic algorithms
Author(s): Billhardt, Holger
Borrajo, Daniel
Maojo, Víctor
Publisher: World Scientific Publishing
Issued date: Feb-2003
Citation: International journal of uncertainty, fuzziness and knowledge-based systems, 2003, vol. 11, n. 1, p. 87-114
URI: http://hdl.handle.net/10016/6744
ISSN: 1793-6411 (online)
0218-4885 (print)
DOI: http://dx.doi.org/10.1142/S0218488503001965
Abstract: The goal of information retrieval (IR) is to provide models and systems that help users to identify the relevant documents to their information needs. Extensive research has been carried out to develop retrieval methods that solve this goal. These IR techniques range from purely syntax-based, considering only frequencies of words, to more semantics-aware approaches. However, it seems clear that there is no single method that works equally well on all collections and for all queries. Prior work suggests that combining the evidence from multiple retrieval experts can achieve significant improvements in retrieval effectiveness. A common problem of expert combination approaches is the selection of both the experts to be combined and the combination function. In most studies the experts are selected from a rather small set of candidates using some heuristics. Thus, only a reduced number of possible combinations is considered and other possibly better solutions are left out. In this paper we propose the use of genetic algorithms to find a suboptimal combination of experts for a document collection at hand. Our approach automatically determines both the experts to be combined and the parameters of the combination function. Because we learn this combination for each specific document collection, this approach allows us to automatically adjust the IR system to specific user needs. To learn retrieval strategies that generalize well on new queries we propose a fitness function that is based on the statistical significance of the average precision obtained on a set of training queries. We test and evaluate the approach on four classical text collections. The results show that the learned combination strategies perform better than any of the individual methods and that genetic algorithms provide a viable method to learn expert combinations. The experiments also evaluate the use of a semantic indexing approach, the context vector model, in combination with classical word matching techniques.
Review: PeerReviewed
Publisher version: http://dx.doi.org/10.1142/S0218488503001965
Keywords: Information retrieval
Data fusion
Genetic algorithms
Context vector model
Rights: © World Scientific Publishing Company
Appears in Collections:DI - PLG - Artículos de Revistas

Refworks Export

SFX Query

Items in E-Archivo are protected by copyright, with all rights reserved, unless otherwise indicated.

 

Valid XHTML 1.0! © Universidad Carlos III de Madrid - Software DSpace - Terms of use - Feedback