Generating Compositional Color Representations from Text

Maheshwari, Paridhi; Jain, Nihal; Vaddamanu, Praneetha; Raut, Dhananjay; Vaishay, Shraiysh; Vinay, Vishwa

Computer Science > Computer Vision and Pattern Recognition

arXiv:2109.10477 (cs)

[Submitted on 22 Sep 2021]

Title:Generating Compositional Color Representations from Text

Authors:Paridhi Maheshwari, Nihal Jain, Praneetha Vaddamanu, Dhananjay Raut, Shraiysh Vaishay, Vishwa Vinay

View PDF

Abstract:We consider the cross-modal task of producing color representations for text phrases. Motivated by the fact that a significant fraction of user queries on an image search engine follow an (attribute, object) structure, we propose a generative adversarial network that generates color profiles for such bigrams. We design our pipeline to learn composition - the ability to combine seen attributes and objects to unseen pairs. We propose a novel dataset curation pipeline from existing public sources. We describe how a set of phrases of interest can be compiled using a graph propagation technique, and then mapped to images. While this dataset is specialized for our investigations on color, the method can be extended to other visual dimensions where composition is of interest. We provide detailed ablation studies that test the behavior of our GAN architecture with loss functions from the contrastive learning literature. We show that the generative model achieves lower Frechet Inception Distance than discriminative ones, and therefore predicts color profiles that better match those from real images. Finally, we demonstrate improved performance in image retrieval and classification, indicating the crucial role that color plays in these downstream tasks.

Comments:	Accepted as a full paper at CIKM 2021
Subjects:	Computer Vision and Pattern Recognition (cs.CV); Information Retrieval (cs.IR); Machine Learning (cs.LG)
Cite as:	arXiv:2109.10477 [cs.CV]
	(or arXiv:2109.10477v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2109.10477

Submission history

From: Paridhi Maheshwari [view email]
[v1] Wed, 22 Sep 2021 01:37:13 UTC (18,182 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Generating Compositional Color Representations from Text

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Generating Compositional Color Representations from Text

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators