Thumbnail
Access Restriction
Subscribed

Author Salton, G. ♦ Yu, C. T.
Source ACM Digital Library
Content type Text
Publisher Association for Computing Machinery (ACM)
File Format PDF
Language English
Subject Keyword Term accuracy ♦ Automatic indexing ♦ Information retrieval ♦ Thesaurus and phrase transformations ♦ Frequency weighting ♦ Content analysis
Abstract The performance of information retrieval systems can be evaluated in a number of different ways. Much of the published evaluation work is based on measuring the retrieval performance of an average user query. Unfortunately, formal proofs are difficult to construct for the average case. In the present study, retrieval evaluation is based on optimizing the performance of a specific user query. The concept of query term accuracy is introduced as the probability of occurrence of a query term in the documents relevant to that query. By relating term accuracy to the frequency of occurrence of the term in the documents of a collection it is possible to give formal proofs of the effectiveness with respect to a given user query of a number of automatic indexing systems that have been used successfully in experimental situations. Among these are inverse document frequency weighting, thesaurus construction, and phrase generation.
Description Affiliation: Cornell Univ., Ithaca, NY (Salton, G.) || Univ. of Alberta, Edmonton, Alberta, Canada (Yu, C. T.)
Age Range 18 to 22 years ♦ above 22 year
Educational Use Research
Education Level UG and PG
Learning Resource Type Article
Publisher Date 2005-08-01
Publisher Place New York
Journal Communications of the ACM (CACM)
Volume Number 20
Issue Number 3
Page Count 8
Starting Page 135
Ending Page 142


Open content in new tab

   Open content in new tab
Source: ACM Digital Library