Robust part-of-speech tagging using a hidden Markov model
Explore this paper's citation graph
Summary
A system for part-of-speech tagging is described, based on a hidden Markov model which can be trained using a corpus of untagged text, which results in a model that correctly tags approximately 96% of the text.
- Type
- article
- Published
- 1992-07-01
- Cited by
- 518
- References
- 20
- OpenAlex
- https://openalex.org/W2100796029
- Semantic Scholar
- https://api.semanticscholar.org/CorpusID:62680996
Keywords
Computer science, Hidden Markov model, Robustness (evolution), Suffix, Artificial intelligence
References
- Three probabilistic language models for a large-vocabulary speech recognizer
- POST: Using Probabilities in Language Processing
- Grammatical Category Disambiguation by Statistical Optimization
- An inequality and associated maximization technique in statistical estimation of probabilistic functions of a Markov process
- A study of English word category prediction based on neutral networks
- Interpolated estimation of Markov source parameters from sparse data
- Statistical language modeling using a small corpus from an application domain
- Tagging text with a probabilistic model
- Finite-State Parsing and Disambiguation
- Error bounds for convolutional codes and an asymptotically optimum decoding algorithm
- Syntactic category disambiguation with neural networks
- Constraint Grammar as a Framework for Parsing Running Text
- Three different probabilistic language models: comparison and combination
- Natural Language Modeling for Phoneme-to-Text Transcription
- A Stochastic Parts Program and Noun Phrase Parser for Unrestricted Text
- Augmenting a Hidden Markov Model for Phrase-Dependent Word Tagging
- Probabilistic Models of Short and Long Distance Word Dependencies in Running Text
- A Cache-Based Natural Language Model for Speech Recognition
- Design of a linguistic statistical decoder for the recognition of continuous speech
- An introduction to the application of the theory of probabilistic functions of a Markov process to automatic speech recognition
Cited by
- Machine Learning for Information Extraction from XML marked-up text on the Semantic Web
- Experiments in automatic word class and word sense identification for information retrieval
- Message Passing Algorithms for the Dirichlet Diffusion Tree
- Semantic feature extraction from technical texts with limited human intervention
- Repurposing Theoretical Linguistic Data for Tool Development and Search
- Adjustment Method of Unbalanced Examples based on Expressions Related to an Event
- A trigram part-of-speech tagger for the Apertium free/open-source machine translation platform
- Efficient Stochastic Part-of-Speech Tagging for Hungarian
- Selective Sampling In Natural Language Learning
- Genetic Algorithms in the Brill Tagger : Moving towards language independence
- An Overview of Corpus-Based Statistics-Oriented (CBSO) Techniques for Natural Language Processing
- Text-to-Speech Conversion of Standard Malay
- A Re-estimation Method for Stochastic Language Modeling from Ambiguous Observations
- Intersection Optimization is NP Complete
- Training Stochastic Grammars From Unlabelled Text Corpora
- A Persian Part-Of-Speech Tagger Based on Morphological Analysis
- Cooperative unsupervised training of the part-of-speech taggers in a bidirectional machine translation system
- An Ontology Enrichment Method for a Pragmatic Information Extraction System gathering Data on Genetic Interactions
- A Methodology for Exploiting Sophisticated Representations for Classification
- High-quality text-to-speech synthesis : an overview
Related papers
- Improving the robustness with multiple sets of HMMs
- A selection method of speech vocabulary for human-robot speech interaction
- Improved hidden Markov model for speech recognition and POS tagging
- Methods for improving robustness of decision tree in Mandarin speech recognition
- Korean Phoneme Recognition Using duration-dependent 3-State Hidden Markov Model
- The double chain markov model
- Hidden Markov Models and Animal Behaviour
- Hybrid approach to speech recognition using hidden Markov models and Markov chains