Accurate Methods for the Statistics of Surprise and Coincidence
Explore this paper's citation graph
Summary
The basis of a measure based on likelihood ratios that can be applied to the analysis of text is described, and in cases where traditional contingency table methods work well, the likelihood ratio tests described here are nearly identical.
- Type
- article
- Published
- 1993-03-01
- Cited by
- 3,034
- References
- 14
- OpenAlex
- https://openalex.org/W2998237634
- Semantic Scholar
- https://api.semanticscholar.org/CorpusID:6465096
Keywords
Surprise, Coincidence, Computer science, Summary statistics, Statistics
References
- A Program for Aligning Sentences in Bilingual Corpora
- Using pathfinder to extract semantic information from text
- Introduction to Modern Information Retrieval
- Introduction To The Theory Of Statistics
- Parsing, Word Associations and Typical Predicate-Argument Relations
- Using latent semantic analysis to improve access to textual information
- A Statistical Approach to Machine Translation
- Pathfinder associative networks: studies in knowledge organization
- Distribution-Free Statistical Tests
- Identifying word correspondence in parallel texts
- An Introduction to the Theory of Statistics
- Identifying Word Correspondences in Parallel Texts
Cited by
- Distributional semantic phrases vs. semantic distributional nonsense: Adjective Modification in Compositional Distributional Semantics
- Evaluation of Web-based Corpora: Effects of Seed Selection and Time Interval
- Uncertainty Management in Information Systems
- Statistical modeling of multiword expressions
- Text mining and IRT for psychiatric and psychological assessment
- esTaBlecimienTo de caTeGorías de análisis linGüísTico a ParTir del discurso de un PacienTe con TrasTorno oBsesivo comPulsivo, medianTe Técnicas de ProcesamienTo auTomáTico del lenGuaje naTural
- Deixis and Fictional Minds
- NLP-based Ontology Learning from Legal Texts. A Case Study
- Acquisition of Qualia Elements from Corpora - Evaluation of a Symbolic Learning Method
- The Israeli-Palestinian Conflict in American, Arab, and British Media: Corpus-Based Critical Discourse Analysis
- Évaluation du potentiel terminologique de candidats termes
- Analysis and construction of noun hypernym hierarchies to enhance Roget's Thesaurus
- Paraphrase Alignment for Synonym Evidence Discovery
- Bootstrapping Biomedical Ontologies for Scientific Text using NELL
- Building and Using Comparable Corpora for Domain-Specific Bilingual Lexicon Extraction
- Automatic Extraction of Subcategorization Frames from the Bulgarian Tree Bank
- Anchors in Context: A corpus analysis of web pages authoring conventions
- Effective and Efficient Correlation Analysis with Application to Market Basket Analysis and Network Community Detection.
- A Workbench for Information Retrieval Experimentation
- Exploring and Visualizing Variation in Language Resources
Related papers
- Book Reviews: Foundations of Statistical Natural Language Processing
- Word Association Norms, Mutual Information, and Lexicography
- Choosing statistical tests: part 12 of a series on evaluation of scientific publications.
- MISUSAGE OF STATISTICS IN MEDICAL RESEARCH
- Problems and mistakes in statistical analysis
- Research ABC : Statistical errors in medical research
- [On the use of mathematical statistics methods in clinical and experimental studies.]
- Statistical Approaches-Reply
- The importance of statistical tools in research
- Does bad inference drive out good?