Designovel's System Description For Fashion-iq Challenge 2019
2019 Β· Jianri Li, Jae-Whan Lee, Woo-Sang Song, et al.
Abstract
This paper describes Designovel's systems which are submitted to the Fashion IQ Challenge 2019. Goal of the challenge is building an image retrieval system where input query is a candidate image plus two text phrases describe user's feedback about visual differences between the candidate image and the search target. We built the systems by combining methods from recent work on deep metric learning, multi-modal retrieval and natual language processing. First, we encode both candidate and target images with CNNs into high-level representations, and encode text descriptions to a single text vector using Transformer-based encoder. Then we compose candidate image vector and text representation into a single vector which is exptected to be biased toward target image vector. Finally, we compute cosine similarities between composed vector and encoded vectors of whole dataset, and rank them in desceding order to get ranked list. We experimented with Fashion IQ 2019 dataset in various settings o
Authors
(none)
Tags
Stats
Related papers
- Fashion IQ: A New Dataset Towards Retrieving Images By Natural Language Feedback (2019)17.43
- Training And Challenging Models For Text-guided Fashion Image Retrieval (2022)0.00
- Searching For Apparel Products From Images In The Wild (2019)0.00
- A Hybrid Multimodal Deep Learning Framework For Intelligent Fashion Recommendation (2025)0.00
- Self-distilled Dynamic Fusion Network For Language-based Fashion Retrieval (2024)5.84
- Deepstyle: Multimodal Search Engine For Fashion And Interior Design (2018)13.17
- Lrvs-fashion: Extending Visual Search With Referring Instructions (2023)0.00
- Fad-vlp: Fashion Vision-and-language Pre-training Towards Unified Retrieval And Captioning (2022)7.81