We gratefully acknowledge support from
the Simons Foundation and member institutions.
Full-text links:

Download:

Current browse context:

cs.CV

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

(what is this?)
CiteULike logo BibSonomy logo Mendeley logo del.icio.us logo Digg logo Reddit logo

Computer Science > Computer Vision and Pattern Recognition

Title: Exploiting Modern Hardware for High-Dimensional Nearest Neighbor Search

Authors: Fabien André
Abstract: Many multimedia information retrieval or machine learning problems require efficient high-dimensional nearest neighbor search techniques. For instance, multimedia objects (images, music or videos) can be represented by high-dimensional feature vectors. Finding two similar multimedia objects then comes down to finding two objects that have similar feature vectors. In the current context of mass use of social networks, large scale multimedia databases or large scale machine learning applications are more and more common, calling for efficient nearest neighbor search approaches.
This thesis builds on product quantization, an efficient nearest neighbor search technique that compresses high-dimensional vectors into short codes. This makes it possible to store very large databases entirely in RAM, enabling low response times. We propose several contributions that exploit the capabilities of modern CPUs, especially SIMD and the cache hierarchy, to further decrease response times offered by product quantization.
Comments: PhD Thesis, 123 Pages
Subjects: Computer Vision and Pattern Recognition (cs.CV); Databases (cs.DB); Information Retrieval (cs.IR); Multimedia (cs.MM); Performance (cs.PF)
Cite as: arXiv:1712.02912 [cs.CV]
  (or arXiv:1712.02912v1 [cs.CV] for this version)

Submission history

From: Fabien André [view email]
[v1] Fri, 8 Dec 2017 02:14:17 GMT (2623kb,D)

Link back to: arXiv, form interface, contact.