Hashing-basierte Indizierung: Anwendungsszenarien, Theorie und Methoden
Explore this paper's citation graph
Summary
Eine Analyse dieser Art ist neu; sie zeigt das enorme Potenzial masgeschneiderter hashing-basierter Indizierungsmethoden wie zum Beispiel dem FuzzyFingerprinting.
- Type
- article
- Published
- 2006-01-01
- Cited by
- 3
- References
- 17
- Access
- Open access
- OpenAlex
- https://openalex.org/W84704405
- Semantic Scholar
- https://api.semanticscholar.org/CorpusID:16890158
Keywords
Humanities, Computer science, Art
References
- Keyword extraction from a single document using word co-occurrence statistical information
- The Reuters Corpus Volume 1 -from Yesterday’s News to Tomorrow’s Language Resources
- Similarity Search in High Dimensions via Hashing
- A Quantitative Analysis and Performance Study for Similarity-Search Methods in High-Dimensional Spaces
- Similarity estimation techniques from rounding algorithms
- Web document clustering: a feasibility demonstration
- Stable distributions, pseudorandom generators, embeddings and data stream computation
- LSH forest: self-tuning indexes for similarity search
- Locality-sensitive hashing scheme based on p-stable distributions
- Approximate Nearest Neighbor: Towards Removing the Curse of Dimensionality
- Managing gigabytes (2nd ed.): compressing and indexing documents and images
- Approximate nearest neighbors
- Fuzzy-Fingerprints for Text-Based Information Retrieval
- Stable Distributions. Models for Heavy Tailed Data
- Near Similarity Search and Plagiarism Analysis
- Identifying and Filtering Near-Duplicate Documents
Cited by
- Technologies for Reusing Text from the Web
- Artificial Intelligence and Machine Learning for Digital Pathology: State-of-the-Art and Future Challenges
- Selecting the right search term in query-based systems for deduplication
- Extension of the Identity Management System Mainzelliste to Reduce Runtimes for Patient Registration in Large Datasets
Related papers
No related papers recorded.