Multi-Cell Multi-Task Convolutional Neural Networks for Diabetic Retinopathy Grading

Zhou, Kang; Gu, Zaiwang; Liu, Wen; Luo, Weixin; Cheng, Jun; Gao, Shenghua; Liu, Jiang

Computer Science > Computer Vision and Pattern Recognition

arXiv:1808.10564 (cs)

[Submitted on 31 Aug 2018 (v1), last revised 11 Oct 2018 (this version, v2)]

Title:Multi-Cell Multi-Task Convolutional Neural Networks for Diabetic Retinopathy Grading

Authors:Kang Zhou, Zaiwang Gu, Wen Liu, Weixin Luo, Jun Cheng, Shenghua Gao, Jiang Liu

View PDF

Abstract:Diabetic Retinopathy (DR) is a non-negligible eye disease among patients with Diabetes Mellitus, and automatic retinal image analysis algorithm for the DR screening is in high demand. Considering the resolution of retinal image is very high, where small pathological tissues can be detected only with large resolution image and large local receptive field are required to identify those late stage disease, but directly training a neural network with very deep architecture and high resolution image is both time computational expensive and difficult because of gradient vanishing/exploding problem, we propose a \textbf{Multi-Cell} architecture which gradually increases the depth of deep neural network and the resolution of input image, which both boosts the training time but also improves the classification accuracy. Further, considering the different stages of DR actually progress gradually, which means the labels of different stages are related. To considering the relationships of images with different stages, we propose a \textbf{Multi-Task} learning strategy which predicts the label with both classification and regression. Experimental results on the Kaggle dataset show that our method achieves a Kappa of 0.841 on test set which is the 4-th rank of all state-of-the-arts methods. Further, our Multi-Cell Multi-Task Convolutional Neural Networks (M$^2$CNN) solution is a general framework, which can be readily integrated with many other deep neural network architectures.

Comments:	Accepted by EMBC 2018
Subjects:	Computer Vision and Pattern Recognition (cs.CV); Image and Video Processing (eess.IV)
Cite as:	arXiv:1808.10564 [cs.CV]
	(or arXiv:1808.10564v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.1808.10564

Submission history

From: Kang Zhou [view email]
[v1] Fri, 31 Aug 2018 01:21:46 UTC (980 KB)
[v2] Thu, 11 Oct 2018 06:10:57 UTC (980 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Multi-Cell Multi-Task Convolutional Neural Networks for Diabetic Retinopathy Grading

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Multi-Cell Multi-Task Convolutional Neural Networks for Diabetic Retinopathy Grading

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators