Effectively Leveraging Attributes for Visual Similarity

Mishra, Samarth; Zhang, Zhongping; Shen, Yuan; Kumar, Ranjitha; Saligrama, Venkatesh; Plummer, Bryan

Computer Science > Computer Vision and Pattern Recognition

arXiv:2105.01695 (cs)

[Submitted on 4 May 2021 (v1), last revised 20 Aug 2021 (this version, v2)]

Title:Effectively Leveraging Attributes for Visual Similarity

Authors:Samarth Mishra, Zhongping Zhang, Yuan Shen, Ranjitha Kumar, Venkatesh Saligrama, Bryan Plummer

View PDF

Abstract:Measuring similarity between two images often requires performing complex reasoning along different axes (e.g., color, texture, or shape). Insights into what might be important for measuring similarity can can be provided by annotated attributes, but prior work tends to view these annotations as complete, resulting in them using a simplistic approach of predicting attributes on single images, which are, in turn, used to measure similarity. However, it is impractical for a dataset to fully annotate every attribute that may be important. Thus, only representing images based on these incomplete annotations may miss out on key information. To address this issue, we propose the Pairwise Attribute-informed similarity Network (PAN), which breaks similarity learning into capturing similarity conditions and relevance scores from a joint representation of two images. This enables our model to identify that two images contain the same attribute, but can have it deemed irrelevant (e.g., due to fine-grained differences between them) and ignored for measuring similarity between the two images. Notably, while prior methods of using attribute annotations are often unable to outperform prior art, PAN obtains a 4-9% improvement on compatibility prediction between clothing items on Polyvore Outfits, a 5% gain on few shot classification of images using Caltech-UCSD Birds (CUB), and over 1% boost to Recall@1 on In-Shop Clothes Retrieval. Implementation available at this https URL

Comments:	Accepted to ICCV 2021
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2105.01695 [cs.CV]
	(or arXiv:2105.01695v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2105.01695

Submission history

From: Samarth Mishra [view email]
[v1] Tue, 4 May 2021 18:28:35 UTC (2,502 KB)
[v2] Fri, 20 Aug 2021 13:48:47 UTC (1,068 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Effectively Leveraging Attributes for Visual Similarity

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Effectively Leveraging Attributes for Visual Similarity

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators