| HOME | HELP | FEEDBACK | SUBSCRIPTIONS | ARCHIVE | SEARCH |
| ||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||
Submitted on February 12, 2006
Accepted on June 12, 2006
Affiliation of the authors: 1 Department of Medical Informatics & Clinical Epidemiology Oregon Health & Science University, Portland, OR, USA ; 2 Medical Informatics Service, University & Hospitals of Geneva, Geneva, Switzerland
* To whom correspondence should be addressed.
Objective Develop and analyze results from an image retrieval test collection.
Methods After participating research groups obtained and assessed results from their systems in the image retrieval task of Cross-Language Evaluation Forum, we assessed the results for common themes and trends. In addition to overall performance, results were analyzed on the basis of topic categories (those most amenable to visual, textual, or mixed approaches) and run categories (those employing queries entered by automated or manual means as well as those using visual, textual, or mixed indexing and retrieval methods). We also assessed results on the different topics and compared the impact of duplicate relevance judgments.
Results A total of 13 research groups participated. Analysis was limited to the best run submitted by each group in each run category. The best results were obtained by systems that combined visual and textual methods. There was substantial variation in performance across topics. Systems employing textual methods were more resilient to visually oriented topics than those using visual methods were to textually oriented topics. The primary performance measure of mean average precision (MAP) was not necessarily associated with other measures, including those possibly more pertinent to real users, such as precision at 10 or 30 images.
Conclusions We developed a test collection amenable to assessing visual and textual methods for image retrieval. Future work must focus on how varying topic and run types affect retrieval performance. Users studies also are necessary to determine the best measures for evaluating the efficacy of image retrieval systems.
This article has been cited by other articles:
![]() |
C. E. Kahn Jr Multilingual Retrieval of Radiology Images RadioGraphics, November 19, 2008; (2008) 291085075. [Abstract] [Full Text] |
||||
![]() |
O. Uzuner, I. Goldstein, Y. Luo, and I. Kohane Identifying Patient Smoking Status from Medical Discharge Records J. Am. Med. Inform. Assoc., January 1, 2008; 15(1): 14 - 24. [Abstract] [Full Text] [PDF] |
||||
| HOME | HELP | FEEDBACK | SUBSCRIPTIONS | ARCHIVE | SEARCH |