Image Matters: Scalable Detection of Offensive and Non-Compliant Content / Logo in Product Images

Gandhi, Shreyansh; Kokkula, Samrat; Chaudhuri, Abon; Magnani, Alessandro; Stanley, Theban; Ahmadi, Behzad; Kandaswamy, Venkatesh; Ovenc, Omer; Mannor, Shie

Computer Science > Computer Vision and Pattern Recognition

arXiv:1905.02234 (cs)

[Submitted on 6 May 2019 (v1), last revised 2 Aug 2019 (this version, v2)]

Title:Image Matters: Scalable Detection of Offensive and Non-Compliant Content / Logo in Product Images

Authors:Shreyansh Gandhi, Samrat Kokkula, Abon Chaudhuri, Alessandro Magnani, Theban Stanley, Behzad Ahmadi, Venkatesh Kandaswamy, Omer Ovenc, Shie Mannor

View PDF

Abstract:In e-commerce, product content, especially product images have a significant influence on a customer's journey from product discovery to evaluation and finally, purchase decision. Since many e-commerce retailers sell items from other third-party marketplace sellers besides their own, the content published by both internal and external content creators needs to be monitored and enriched, wherever possible. Despite guidelines and warnings, product listings that contain offensive and non-compliant images continue to enter catalogs. Offensive and non-compliant content can include a wide range of objects, logos, and banners conveying violent, sexually explicit, racist, or promotional messages. Such images can severely damage the customer experience, lead to legal issues, and erode the company brand. In this paper, we present a computer vision driven offensive and non-compliant image detection system for extremely large image datasets. This paper delves into the unique challenges of applying deep learning to real-world product image data from retail world. We demonstrate how we resolve a number of technical challenges such as lack of training data, severe class imbalance, fine-grained class definitions etc. using a number of practical yet unique technical strategies. Our system combines state-of-the-art image classification and object detection techniques with budgeted crowdsourcing to develop a solution customized for a massive, diverse, and constantly evolving product catalog.

Comments:	10 pages
Subjects:	Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
Cite as:	arXiv:1905.02234 [cs.CV]
	(or arXiv:1905.02234v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.1905.02234

Submission history

From: Shreyansh Gandhi [view email]
[v1] Mon, 6 May 2019 18:35:28 UTC (11,153 KB)
[v2] Fri, 2 Aug 2019 07:38:26 UTC (10,847 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Image Matters: Scalable Detection of Offensive and Non-Compliant Content / Logo in Product Images

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Image Matters: Scalable Detection of Offensive and Non-Compliant Content / Logo in Product Images

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators