Learning to construct knowledge bases from the World Wide Web
Explore this paper's citation graph
Summary
The goal of the research described here is to automatically create a computer understandable knowledge base whose content mirrors that of the World Wide Web, and several machine learning algorithms for this task are described, and promising initial results with a prototype system that has created a knowledge base describing university people, courses, and research projects.
- Type
- article
- Published
- 2000-04-01
- Cited by
- 560
- References
- 70
- Access
- Open access
- OpenAlex
- https://openalex.org/W2033709196
- Semantic Scholar
- https://api.semanticscholar.org/CorpusID:5303928
Keywords
Computer science, Knowledge base, Hypertext, Hyperlink, Construct (python library)
References
- The fourth text REtrieval conference
- Information Theory: 1948-1998 - Guest Editorial
- Estimating Probabilities: A Crucial Task in Machine Learning
- Acquisition of Linguistic Patterns for Knowledge-based Information Extraction
- Improving Text Classification by Shrinkage in a Hierarchy of Classes
- Learning text analysis rules for domain-specific natural language processing
- A comparison of event models for naive bayes text classification
- Wrapper Induction for Information Extraction
- Multistrategy Learning for Information Extraction
- Syskill & Webert: Identifying Interesting Web Sites
- Category Translation: Learning to Understand Information on the Internet
- Connectionist, Statistical and Symbolic Approaches to Learning for Natural Language Processing
- Hierarchically Classifying Documents Using Very Few Words
- A Case Study in Using Linguistic Phrases for Text Categorization on the WWW
- Toward Optimal Feature Selection
- Optimized rule induction
- An Empirical Study of Automated Dictionary Construction for Information Extraction in Three Domains
- Classifying news stories using memory based reasoning
- Web Watcher: A Tour Guide for the World Wide Web
- Towards language independent automated learning of text categorization models
Cited by
- Learning to extract information from large websites using sequential models
- Mining and Re-ranking for Answering Biographical Queries on the Web
- USING CONCEPT HIERARCHIES TO ENHANCE USER QUERIES IN WEB-BASED INFORMATION RETRIEVAL
- Création de surcouche de documents hypertextes et traitement du langage naturel
- Learning Text Extraction Rules, without Ignoring Stop Words
- Use of Ontologies for Cross-lingual Information Management in the Web
- Using Support Vector Machines for Classifying Large Sets of Multi-Represented Objects
- Querying and extracting heterogeneous graphs from structured data and unstrutured content
- Wissensbasiertes Text-Mining mit SynDiKATe
- Searching the Web: learning based techniques
- Learning of Ontologies from the Web: the Analysis of Existent Approaches
- Generating Dynamic and Adaptive Knowledge Models for Web-Based Resources
- A Semantic Memory for Incremental Ontology Population
- Effects of a professional development initiative on technology innovation in the elementary school
- What You Seek Is What You Get: Extraction of Class Attributes from Query Logs
- Clustering Presentation of Web Image Retrieval Results using Textual Information and Image Features
- Mining Domain Specific Texts and Glossaries to Evaluate and Enrich Domain Ontologies
- Towards a Universal Web Wrapper
- Automated Population of Cyc: Extracting Information about Named-entities from the Web
- Object Matching for Information Integration: A Profiler-Based Approach
Related papers
- AN IMPLEMENTATION OF WEB PERSONALIZATION USING WEB MINING TECHNIQUES
- Web mining techniques for recommendation and personalization
- Investigation of Heterogeneous Approach to Fact Invention of Web Users’ Web Access Behaviour
- A Novel Overview and Evolution of World Wide Web: Comparison from Web 1.0 to Web 3.0
- Web Personalization With Web Usage Mining Technics and Association Rules
- Knowledge Discovery from Web Data for Web Personalization
- Evolution of the weband E-Learning application
- A SYSTEMATIC REVIEW ON DATA PREPROCESSING AND PATTERN DISCOVERY OF WEB USAGE MINING