跳到主要导航 跳到搜索 跳到主要内容

C-approximate nearest neighbor query algorithm based on learning for high-dimensional data

  • Fudan University

科研成果: 期刊稿件文章同行评审

摘要

Under the filter-and-refine framework and based on the learning techniques, a data-aware method for c-approximate nearest neighbor query for high-dimensional data is proposed in this paper. The study claims that data after random projection satisfies the entropy-maximizing criterion which is needed by the semantic hashing. The binary codes after random projection are treated as the labels, and a group of classifiers are trained, which are used for predicting the binary code for the query. The data objects are selected who's Hamming distances between the query satisfying the threshold as the candidates. The real distances are evaluated on the candidate subset and the smallest one is returned. Experimental results on the synthetic datasets and the real datasets show that this method outperforms the existing work with shorter binary code, in addition, the performance and the result quality can be easily tuned.

源语言英语
页(从-至)2018-2031
页数14
期刊Ruan Jian Xue Bao/Journal of Software
23
8
DOI
出版状态已出版 - 8月 2012

指纹

探究 'C-approximate nearest neighbor query algorithm based on learning for high-dimensional data' 的科研主题。它们共同构成独一无二的指纹。

引用此