Fast and exact fixed-radius neighbor search based on sorting

Author:

Chen Xinye1ORCID,Güttel Stefan2

Affiliation:

1. Charles University Prague, Prague, Czech Republic

2. University of Manchester, Manchester, United Kingdom

Abstract

Fixed-radius near neighbor search is a fundamental data operation that retrieves all data points within a user-specified distance to a query point. There are efficient algorithms that can provide fast approximate query responses, but they often have a very compute-intensive indexing phase and require careful parameter tuning. Therefore, exact brute force and tree-based search methods are still widely used. Here we propose a new fixed-radius near neighbor search method, called SNN, that significantly improves over brute force and tree-based methods in terms of index and query time, provably returns exact results, and requires no parameter tuning. SNN exploits a sorting of the data points by their first principal component to prune the query search space. Further speedup is gained from an efficient implementation using high-level basic linear algebra subprograms (BLAS). We provide theoretical analysis of our method and demonstrate its practical performance when used stand-alone and when applied within the DBSCAN clustering algorithm.

Funder

Royal Society Industry Fellowship

Publisher

PeerJ

Reference64 articles.

1. Refining a k-nearest neighbor graph for a computationally efficient spectral clustering;Alshammari;Pattern Recognition,2021

2. ANN-benchmarks: a benchmarking tool for approximate nearest neighbor algorithms;Aumüller;Information Systems,2020

3. Speeding up the Xbox recommender system using a Euclidean transformation for inner-product spaces;Bachrach,2014

4. LSH Forest: self-tuning indexes for similarity search;Bawa,2005

5. Multidimensional binary search trees used for associative searching;Bentley;Communications of the ACM,1975a

同舟云学术

1.学者识别学者识别

2.学术分析学术分析

3.人才评估人才评估

"同舟云学术"是以全球学者为主线,采集、加工和组织学术论文而形成的新型学术文献查询和分析系统,可以对全球学者进行文献检索和人才价值评估。用户可以通过关注某些学科领域的顶尖人物而持续追踪该领域的学科进展和研究前沿。经过近期的数据扩容,当前同舟云学术共收录了国内外主流学术期刊6万余种,收集的期刊论文及会议论文总量共计约1.5亿篇,并以每天添加12000余篇中外论文的速度递增。我们也可以为用户提供个性化、定制化的学者数据。欢迎来电咨询!咨询电话:010-8811{复制后删除}0370

www.globalauthorid.com

TOP

Copyright © 2019-2024 北京同舟云网络信息技术有限公司
京公网安备11010802033243号  京ICP备18003416号-3