A performance evaluation of similarity measures, document term weighting schemes and representations in a Boolean environment
Explore this paper's citation graph
Summary
It is felt, but does need further examination, that while the absolute effectiveness of any ranking algorithms may vary with the environment, the relative effectiveness of the ranking algorithms will be invariant.
- Type
- article
- Published
- 1980-06-23
- Cited by
- 90
- References
- 19
- OpenAlex
- https://openalex.org/W1991240715
- Semantic Scholar
- https://api.semanticscholar.org/CorpusID:16561941
Keywords
Weighting, Term (time), Computer science, Similarity (geometry), Information retrieval
References
- User directed relevance feedback
- The use of automatically-obtained keyword classifications for information retrieval
- An Evaluation of Factors Affecting Document Ranking by Information Retrieval Systems.
- Stemming as a system design consideration
- Automatic abstracting and indexing—survey and recommendations
- Index term weighting
- A probabilistic approach to automatic keyword indexing. Part I. On the Distribution of Specialty Words in a Technical Literature
- On Relevance, Probabilistic Indexing and Information Retrieval
- A probabilistic approach to automatic keyword indexing. Part II. An algorithm for probabilistic indexing
- A framework for comparing term association measures
- A theoretical basis for the use of co-occurence data in information retrieval
- Automatic ranked output from boolean searches in SIRE
- Theory and Methods of Scaling.
- Numerical Taxonomy: The Principles and Practice of Numerical Classification
- Theory and Methods of Scaling
Cited by
- A Comparison of Statistical Filtering Methods for Automatic Term Extraction for Domain Analysis
- Automated illustration of multimedia stories
- Equivalence topologique entre mesures de proximité
- Keyword extraction from a single document using word co-occurrence statistical information
- Key word extraction from a document using word co-occurrence statistical information
- Non-parametric Information-Theoretic Measures of One-Dimensional Distribution Functions from Continuous Time Series
- A generic software library for creating multimedia browse/search applications
- Development of a flexible tool for the automatic comparison of bibliographic records. Application to sample collections - Développement d"un logiciel flexible pour la comparaison de notices bibliographiques et application à différentes collections
- Context and structure in automated full-text information access
- Information Extraction in the Web Era
- A framework for understanding user interaction with content-based image retrieval: model, interface and users
- Report on the TREC-4 Experiment: Combining Probabilistic and Vector-Space Schemes
- Similarity Measures for Categorical Data: A Comparative Evaluation
- New Primitives for Secure Resource and Computation Verification
- Source Code Retrieval using Conceptual Graphs
- From information retrieval to hypertext and back again: the role of interaction in the information exploration interface
- Unterstützung von Information-Retrieval-Dialogen mit Informationssystemen durch interaktive Informationsvisualisierung
- Using Genetic Algorithm to Improve Information Retrieval Systems
- Supporting data mining of large databases by visual feedback queries
- Comparing Dissimilarity Measures for Content-Based Image Retrieval
Related papers
- Introduction to Modern Information Retrieval
- An algorithm for suffix stripping
- Term-Weighting Approaches in Automatic Text Retrieval
- Information Retrieval: Data Structures and Algorithms
- STUDIES WITH THE ELECTROCARDIOGRAPH ON THE ACTION OF THE VAGUS NERVE ON THE HUMAN HEART
- A statistical interpretation of term specificity and its application in retrieval
- Measures of similarity among fuzzy concepts: A comparative analysis
- Improving retrieval performance by relevance feedback
- The SMART Retrieval System—Experiments in Automatic Document Processing
- Computational analysis of present-day American English