Fuzzy term proximity information retrieval model and system

The huge size of digital data accentuates the scientific challenge of information retrieval (IR) consisting in finding a compromise between recall and precision. We propose an IR model based on fuzzy proximity (FP) of the query terms which is aimed to high precision. It combines the expressivity of the Boolean query model and the ranking of the documents thanks to the use of proximity. Each keyword defines an influence zone at the query evaluation time. The fuzzy operations associated to the traditional Boolean operators propagate the proximity to the root of the query tree. The FP model was largely validated on the traditional test collections and at the 2005 and 2006 editions of the international IR evaluation campaigns (TREC, CLEF and INEX 2006). The results obtained with the automatically built queries are equivalent to the baselines (Okapi/Lucy and vector/MG). Moreover, with manual queries adapted to FP, the results are better than the baselines.

Data and Resources

Additional Info

Field Value
Source https://theses.hal.science/tel-00785143
Author Mercier, Annabelle
Maintainer CCSD
Last Updated May 14, 2026, 16:31 (UTC)
Created May 14, 2026, 16:31 (UTC)
Identifier NNT: 2006EMSE0024
Language fr
Rights https://about.hal.science/hal-authorisation-v1/
contributor Département Réseaux, Information, Multimédia (RIM-ENSMSE) ; École des Mines de Saint-Étienne (Mines Saint-Étienne MSE) ; Institut Mines-Télécom [Paris] (IMT)-Institut Mines-Télécom [Paris] (IMT)-Centre G2I
creator Mercier, Annabelle
date 2006-11-13T00:00:00
harvest_object_id c5f44f3f-6289-4f1c-9e31-c19939bd2797
harvest_source_id 3374d638-d20b-4672-ba96-a23232d55657
harvest_source_title test moissonnage SELUNE
metadata_modified 2026-01-19T00:00:00
set_spec type:THESE