Results 51 to 60 of about 325,256 (164)
Multivariate Similarity Search - A Call for a New Breed of Similarity Search Algorithms
The similarity search task involves identifying pairs of similar vectors, e.g., time series. For example, given a query q, the user might wish to find all vectors in a dataset with a cosine similarity with q higher than a threshold t, or to find the top-k most similar vectors with q, using Euclidean distance.
Odysseas Papapetrou, Jens E. d'Hondt
openaire +1 more source
Lightweight Semantic Similarity Search of Data Products
The increasing popularity of digital services provided by large enterprises has determined a proliferation of data products, determining several data management issues.
Antonio D'Ambrosio +5 more
doaj +1 more source
Background There is an increasing amount of unstructured medical data that can be analysed for different purposes. However, information extraction from free text data may be particularly inefficient in the presence of spelling errors. Existing approaches
Hegler Tissot, Richard Dobson
doaj +1 more source
APOTHEOSIS: An efficient approximate similarity search system
APOTHEOSIS is a tool for efficiently identifying and comparing data similarity in large datasets, addressing challenges faced by traditional methods such as scalability and speed.
Daniel Huici +2 more
doaj +1 more source
Gaussian Kernel-Based LSH for High-Dimensional Similarity Search
High-dimensional similarity search remains a critical challenge in machine learning, particularly when data lie on complex, non-linear manifolds that undermine the effectiveness of classical Locality-Sensitive Hashing (LSH). This work introduces Gaussian
Masrat Rasool +2 more
doaj +1 more source
Identification of a Putative α-synuclein Radioligand Using an in silico Similarity Search. [PDF]
Janssen B +17 more
europepmc +1 more source
From Fact Drafts to Operational Systems: Semantic Search in Legal Decisions Using Fact Drafts
This research paper presents findings from an investigation in the semantic similarity search task within the legal domain, using a corpus of 1172 Hungarian court decisions.
Gergely Márk Csányi +6 more
doaj +1 more source
Iterative machine learning-based chemical similarity search to identify novel chemical inhibitors. [PDF]
Durai P, Lee SJ, Lee JW, Pan CH, Park K.
europepmc +1 more source
SIMBSIG: similarity search and clustering for biobank-scale data. [PDF]
Adamer MF +3 more
europepmc +1 more source
Billion-Scale Similarity Search Using a Hybrid Indexing Approach with Advanced Filtering
This paper presents a novel approach for similarity search with complex filtering capabilities on billion-scale datasets, optimized for CPU inference. Our method extends the classical IVF-Flat index structure to integrate multi-dimensional filters.
Emanuilov Simeon, Dimov Aleksandar
doaj +1 more source

