A statistical interpretation of term specificity and its application in retrieval
Karen Spärck Jones
University of Cambridge
阅读操作
确认中在文库中上传 PDF 后可生成中文音频讲解。
摘要与影响
The exhaustivity of document descriptions and the specificity of index terms are usually regarded as independent. It is suggested that specificity should be interpreted statistically, as a function of term use rather than of term meaning. The effects on retrieval of variations in term specificity are examined, experiments with three test collections showing, in particular, that frequently‐occurring terms are required for good overall performance. It is argued that terms should be weighted according to collection frequency, so that matches on less frequent, more specific, terms are of greater value than matches on frequent terms. Results for the test collections show that considerable improvements in performance are obtained with this very simple procedure.
逐年被引趋势
关键指标
同类平均 = 1
同领域 · 同年份 · 同类型
Google Scholar 与 OpenAlex 的被引统计范围不同,数值存在差异属正常。
AI 辅助阅读
依据:摘要
可就本文提问;依据不足时会说明。
学术脉络
学科主题
计算机 / AIAdvanced Text Analysis Techniques
Semantic Web and Ontologies · Natural Language Processing Techniques
参考文献 41
此处列出前 3 条
引用本文 502
按被引量排序,此处列出前 3 条