FFAVOD: Feature Fusion Architecture for Video Object Detection

Perreault, Hughes; Bilodeau, Guillaume-Alexandre; Saunier, Nicolas; Héritier, Maguelonne

Computer Science > Computer Vision and Pattern Recognition

arXiv:2109.07298 (cs)

[Submitted on 15 Sep 2021]

Title:FFAVOD: Feature Fusion Architecture for Video Object Detection

Authors:Hughes Perreault, Guillaume-Alexandre Bilodeau, Nicolas Saunier, Maguelonne Héritier

View PDF

Abstract:A significant amount of redundancy exists between consecutive frames of a video. Object detectors typically produce detections for one image at a time, without any capabilities for taking advantage of this redundancy. Meanwhile, many applications for object detection work with videos, including intelligent transportation systems, advanced driver assistance systems and video surveillance. Our work aims at taking advantage of the similarity between video frames to produce better detections. We propose FFAVOD, standing for feature fusion architecture for video object detection. We first introduce a novel video object detection architecture that allows a network to share feature maps between nearby frames. Second, we propose a feature fusion module that learns to merge feature maps to enhance them. We show that using the proposed architecture and the fusion module can improve the performance of three base object detectors on two object detection benchmarks containing sequences of moving road users. Additionally, to further increase performance, we propose an improvement to the SpotNet attention module. Using our architecture on the improved SpotNet detector, we obtain the state-of-the-art performance on the UA-DETRAC public benchmark as well as on the UAVDT dataset. Code is available at this https URL.

Comments:	Accepted for publication in Pattern Recognition Letters
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2109.07298 [cs.CV]
	(or arXiv:2109.07298v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2109.07298

Submission history

From: Hughes Perreault [view email]
[v1] Wed, 15 Sep 2021 13:53:21 UTC (288 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:FFAVOD: Feature Fusion Architecture for Video Object Detection

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:FFAVOD: Feature Fusion Architecture for Video Object Detection

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators