SLOSH: Set LOcality Sensitive Hashing via Sliced-Wasserstein Embeddings

Lu, Yuzhe; Liu, Xinran; Soltoggio, Andrea; Kolouri, Soheil

Computer Science > Machine Learning

arXiv:2112.05872 (cs)

[Submitted on 11 Dec 2021 (v1), last revised 8 Feb 2022 (this version, v2)]

Title:SLOSH: Set LOcality Sensitive Hashing via Sliced-Wasserstein Embeddings

Authors:Yuzhe Lu, Xinran Liu, Andrea Soltoggio, Soheil Kolouri

View PDF

Abstract:Learning from set-structured data is an essential problem with many applications in machine learning and computer vision. This paper focuses on non-parametric and data-independent learning from set-structured data using approximate nearest neighbor (ANN) solutions, particularly locality-sensitive hashing. We consider the problem of set retrieval from an input set query. Such retrieval problem requires: 1) an efficient mechanism to calculate the distances/dissimilarities between sets, and 2) an appropriate data structure for fast nearest neighbor search. To that end, we propose Sliced-Wasserstein set embedding as a computationally efficient "set-2-vector" mechanism that enables downstream ANN, with theoretical guarantees. The set elements are treated as samples from an unknown underlying distribution, and the Sliced-Wasserstein distance is used to compare sets. We demonstrate the effectiveness of our algorithm, denoted as Set-LOcality Sensitive Hashing (SLOSH), on various set retrieval datasets and compare our proposed embedding with standard set embedding approaches, including Generalized Mean (GeM) embedding/pooling, Featurewise Sort Pooling (FSPool), and Covariance Pooling and show consistent improvement in retrieval results. The code for replicating our results is available here: \href{this https URL}{this https URL}.

Subjects:	Machine Learning (cs.LG); Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2112.05872 [cs.LG]
	(or arXiv:2112.05872v2 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2112.05872

Submission history

From: Soheil Kolouri [view email]
[v1] Sat, 11 Dec 2021 00:10:05 UTC (2,272 KB)
[v2] Tue, 8 Feb 2022 19:03:37 UTC (2,273 KB)

Computer Science > Machine Learning

Title:SLOSH: Set LOcality Sensitive Hashing via Sliced-Wasserstein Embeddings

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:SLOSH: Set LOcality Sensitive Hashing via Sliced-Wasserstein Embeddings

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators