GRIT: Graph-based Recall Improvement for Task-oriented E-commerce Queries

Kulkarni, Hrishikesh; Kallumadi, Surya; MacAvaney, Sean; Goharian, Nazli; Frieder, Ophir

doi:10.1145/3701716.3717859

Computer Science > Information Retrieval

arXiv:2504.05310 (cs)

[Submitted on 16 Feb 2025]

Title:GRIT: Graph-based Recall Improvement for Task-oriented E-commerce Queries

Authors:Hrishikesh Kulkarni, Surya Kallumadi, Sean MacAvaney, Nazli Goharian, Ophir Frieder

View PDF HTML (experimental)

Abstract:Many e-commerce search pipelines have four stages, namely: retrieval, filtering, ranking, and personalized-reranking. The retrieval stage must be efficient and yield high recall because relevant products missed in the first stage cannot be considered in later stages. This is challenging for task-oriented queries (queries with actionable intent) where user requirements are contextually intensive and difficult to understand. To foster research in the domain of e-commerce, we created a novel benchmark for Task-oriented Queries (TQE) by using LLM, which operates over the existing ESCI product search dataset. Furthermore, we propose a novel method 'Graph-based Recall Improvement for Task-oriented queries' (GRIT) to address the most crucial first-stage recall improvement needs. GRIT leads to robust and statistically significant improvements over state-of-the-art lexical, dense, and learned-sparse baselines. Our system supports both traditional and task-oriented e-commerce queries, yielding up to 6.3% recall improvement. In the indexing stage, GRIT first builds a product-product similarity graph using user clicks or manual annotation data. During retrieval, it locates neighbors with higher contextual and action relevance and prioritizes them over the less relevant candidates from the initial retrieval. This leads to a more comprehensive and relevant first-stage result set that improves overall system recall. Overall, GRIT leverages the locality relationships and contextual insights provided by the graph using neighboring nodes to enrich the first-stage retrieval results. We show that the method is not only robust across all introduced parameters, but also works effectively on top of a variety of first-stage retrieval methods.

Comments:	LLM4ECommerce at WWW 2025
Subjects:	Information Retrieval (cs.IR)
Cite as:	arXiv:2504.05310 [cs.IR]
	(or arXiv:2504.05310v1 [cs.IR] for this version)
	https://doi.org/10.48550/arXiv.2504.05310
Journal reference:	Companion Proceedings of the ACM Web Conference 2025 (WWW Companion 25), April 28-May 2, 2025, Sydney, NSW, Australia. ACM, New York, NY, USA, 10 pages
Related DOI:	https://doi.org/10.1145/3701716.3717859

Submission history

From: Hrishikesh Kulkarni [view email]
[v1] Sun, 16 Feb 2025 16:21:49 UTC (1,592 KB)

Computer Science > Information Retrieval

Title:GRIT: Graph-based Recall Improvement for Task-oriented E-commerce Queries

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Information Retrieval

Title:GRIT: Graph-based Recall Improvement for Task-oriented E-commerce Queries

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators