Computer Science > Computer Vision and Pattern Recognition

arXiv:2302.08902 (cs)

[Submitted on 16 Feb 2023 (v1), last revised 7 Mar 2023 (this version, v4)]

Title:Fashion Image Retrieval with Multi-Granular Alignment

Authors:Jinkuan Zhu, Hao Huang, Qiao Deng, Xiyao Li

View PDF

Abstract:Fashion image retrieval task aims to search relevant clothing items of a query image from the gallery. The previous recipes focus on designing different distance-based loss functions, pulling relevant pairs to be close and pushing irrelevant images apart. However, these methods ignore fine-grained features (e.g. neckband, cuff) of clothing images. In this paper, we propose a novel fashion image retrieval method leveraging both global and fine-grained features, dubbed Multi-Granular Alignment (MGA). Specifically, we design a Fine-Granular Aggregator(FGA) to capture and aggregate detailed patterns. Then we propose Attention-based Token Alignment (ATA) to align image features at the multi-granular level in a coarse-to-fine manner. To prove the effectiveness of our proposed method, we conduct experiments on two sub-tasks (In-Shop & Consumer2Shop) of the public fashion datasets DeepFashion. The experimental results show that our MGA outperforms the state-of-the-art methods by 1.8% and 0.6% in the two sub-tasks on the R@1 metric, respectively.

Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2302.08902 [cs.CV]
	(or arXiv:2302.08902v4 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2302.08902

Submission history

From: Jinkuan Zhu [view email]
[v1] Thu, 16 Feb 2023 10:43:31 UTC (1,078 KB)
[v2] Wed, 22 Feb 2023 14:41:35 UTC (1,079 KB)
[v3] Mon, 27 Feb 2023 12:50:16 UTC (1,079 KB)
[v4] Tue, 7 Mar 2023 07:18:22 UTC (1,079 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Fashion Image Retrieval with Multi-Granular Alignment

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Fashion Image Retrieval with Multi-Granular Alignment

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators