SimDex: Exploiting Model Similarity in Exact Matrix Factorization Recommendations

Abuzaid, Firas; Sethi, Geet; Bailis, Peter; Zaharia, Matei

Computer Science > Information Retrieval

arXiv:1706.01449v1 (cs)

[Submitted on 5 Jun 2017 (this version), latest version 15 Mar 2019 (v3)]

Title:SimDex: Exploiting Model Similarity in Exact Matrix Factorization Recommendations

Authors:Firas Abuzaid, Geet Sethi, Peter Bailis, Matei Zaharia

View PDF

Abstract:We present SimDex, a new technique for serving exact top-K recommendations on matrix factorization models that measures and optimizes for the similarity between users in the model. Previous serving techniques presume a high degree of similarity (e.g., L2 or cosine distance) among users and/or items in MF models; however, as we demonstrate, the most accurate models are not guaranteed to exhibit high similarity. As a result, brute-force matrix multiply outperforms recent proposals for top-K serving on several collaborative filtering tasks. Based on this observation, we develop SimDex, a new technique for serving matrix factorization models that automatically optimizes serving based on the degree of similarity between users, and outperforms existing methods in both the high-similarity and low-similarity regimes. SimDexfirst measures the degree of similarity among users via clustering and uses a cost-based optimizer to either construct an index on the model or defer to blocked matrix multiply. It leverages highly efficient linear algebra primitives in both cases to deliver predictions either from its index or from brute-force multiply. Overall, SimDex runs an average of 2x and up to 6x faster than highly optimized baselines for the most accurate models on several popular collaborative filtering datasets.

Comments:	13 pages, 8 figures
Subjects:	Information Retrieval (cs.IR); Databases (cs.DB); Data Structures and Algorithms (cs.DS); Performance (cs.PF)
Cite as:	arXiv:1706.01449 [cs.IR]
	(or arXiv:1706.01449v1 [cs.IR] for this version)
	https://doi.org/10.48550/arXiv.1706.01449

Submission history

From: Firas Abuzaid [view email]
[v1] Mon, 5 Jun 2017 17:56:43 UTC (373 KB)
[v2] Thu, 2 Aug 2018 22:08:15 UTC (1,071 KB)
[v3] Fri, 15 Mar 2019 00:52:25 UTC (1,227 KB)

Computer Science > Information Retrieval

Title:SimDex: Exploiting Model Similarity in Exact Matrix Factorization Recommendations

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Information Retrieval

Title:SimDex: Exploiting Model Similarity in Exact Matrix Factorization Recommendations

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators