Deep Learning Based Large Scale Visual Recommendation And Search For E-commerce
2017 Β· Devashish Shankar, Sujay Narumanchi, H A Ananya, et al.
Abstract
In this paper, we present a unified end-to-end approach to build a large scale Visual Search and Recommendation system for e-commerce. Previous works have targeted these problems in isolation. We believe a more effective and elegant solution could be obtained by tackling them together. We propose a unified Deep Convolutional Neural Network architecture, called VisNet, to learn embeddings to capture the notion of visual similarity, across several semantic granularities. We demonstrate the superiority of our approach for the task of image retrieval, by comparing against the state-of-the-art on the Exact Street2Shop dataset. We then share the design decisions and trade-offs made while deploying the model to power Visual Recommendations across a catalog of 50M products, supporting 2K queries a second at Flipkart, India's largest e-commerce company. The deployment of our solution has yielded a significant business impact, as measured by the conversion-rate.
Authors
(none)
Tags
Stats
Related papers
- Zero-shot Retrieval For Scalable Visual Search In A Two-sided Marketplace (2025)1.57
- Learning A Unified Embedding For Visual Search At Pinterest (2019)10.85
- Retrieving Similar E-commerce Images Using Deep Learning (2019)0.00
- V\(^2\)L: Leveraging Vision And Vision-language Models Into Large-scale Product Retrieval (2022)0.00
- Improving Visual Recommendation On E-commerce Platforms Using Vision-language Models (2025)0.00
- From Pixels To Purchase: Building And Evaluating A Taxonomy-decoupled Visual Search Engine For Home Goods E-commerce (2026)0.00
- Visually Similar Products Retrieval For Shopsy (2022)2.26
- Onevision: An End-to-end Generative Framework For Multi-view E-commerce Vision Search (2025)0.00