Semantic Similarity Based on Corpus Statistics and Lexical Taxonomy
Explore this paper's citation graph
Summary
This paper presents a new approach for measuring semantic similarity/distance between words and concepts that combines a lexical taxonomy structure with corpus statistical information so that the semantic distance between nodes in the semantic space constructed by the taxonomy can be better quantified with the computational evidence derived from a distributional analysis of corpus data.
- Type
- preprint
- Published
- 1997-08-01
- Cited by
- 3,496
- References
- 19
- Access
- Open access
- OpenAlex
- https://openalex.org/W2100935296
- Semantic Scholar
- https://api.semanticscholar.org/CorpusID:1359050
Keywords
Semantic similarity, Computer science, Artificial intelligence, Taxonomy (biology), Natural language processing
References
- WordNet and Distributional Analysis: A Class-based Approach to Lexical Discovery
- Use of syntactic context to produce term association lists for text retrieval
- Nouns in WordNet: A Lexical Inheritance System
- Word sense disambiguation for free-text indexing using a massive semantic network
- A Semantic Concordance
- Information Retrieval Based on Conceptual Distance in is-a Hierarchies
- Similarity between Words Computed by Spreading Activation on an English Dictionary
- Development and application of a metric on semantic nets
- Co-Occurrence Vectors From Corpora vs. Distance Vectors From Dictionaries
- Experiments on using semantic distances between words in image caption retrieval
- Introduction to WordNet: An On-line Lexical Database
- Contextual correlates of semantic similarity
- Noun Classification From Predicate-Argument Structures
- Information Retrieval Using Robust Natural Language Processing
- Word Association Norms, Mutual Information, and Lexicography
- Using Information Content to Evaluate Semantic Similarity in a Taxonomy
- Lexical Cohesion Computed by Thesaural relations as an indicator of the structure of text
- Lexical cohesion computed by thesaural relations as an indicator of the structure of text
- A Proposal for Word Sense Disambiguation using Conceptual Distance
- Using WordNet in a Knowledge-Based Approach to Information Retrieval
Cited by
- UNT: A Supervised Synergistic Approach to Semantic Text Similarity
- Learning to Merge Word Senses
- Découverte et analyse des communautés implicites par une approche sémantique en ligne : l'outil WebTribe
- X-Similarity: Computing Semantic Similarity between Concepts from Different Ontologies
- Semantic Relatedness Using Salient Semantic Analysis
- Statistical modeling of multiword expressions
- Fuzzy Semantic Matching in (Semi-)Structured XML Documents - Indexation of Noisy Documents
- Measuring Semantic Similarity in Short Texts through Greedy Pairing and Word Semantics
- Graph Kernels and Applications in Bioinformatics
- OntoQuest: A Physician Decision Support System based on Ontological Queries of the Hospital Database
- Approches de recherche multimédia dans des documents semi-structurés : utilisation du contexte textuel et structurel pour la sélection d'objets multimédia
- Prior Disambiguation of Word Tensors for Constructing Sentence Vectors
- Towards structured representation of academic search results
- Sentence Similarity based on Relevance
- Bayesian ontology querying for accurate and noise-tolerant semantic searches
- Exploiting disjointness axioms to improve semantic similarity measures
- Determining the difficulty of Word Sense Disambiguation
- Dealing with Acronyms in Biomedical Texts
- Keyword Weight Propagation for Indexing Structured Web Content
- Roget's thesaurus and semantic similarity
Related papers
- Correcting replicate variation in spectroscopic data by machine learning and model-based pre-processing
- To Replicate or Not To Replicate Queries in the Presence of Autonomous Participants
- To Replicate or Not To Replicate
- Mining user similarity from semantic trajectories
- Computing method for semantic similarity of words based on How Net
- A Hybrid Approach for Measuring Semantic Similarity between Documents and its Application in Mining the Knowledge Repositories
- The Research of Word Similarity in Semantic Retrieval