Guided multi-branch learning systems for sound event detection with sound separation

Huang, Yuxin; Lin, Liwei; Ma, Shuo; Wang, Xiangdong; Liu, Hong; Qian, Yueliang; Liu, Min; Ouch, Kazushige

Computer Science > Sound

arXiv:2007.10638 (cs)

[Submitted on 21 Jul 2020 (v1), last revised 2 Nov 2020 (this version, v2)]

Title:Guided multi-branch learning systems for sound event detection with sound separation

Authors:Yuxin Huang, Liwei Lin, Shuo Ma, Xiangdong Wang, Hong Liu, Yueliang Qian, Min Liu, Kazushige Ouch

View PDF

Abstract:In this paper, we describe in detail our systems for DCASE 2020 Task 4. The systems are based on the 1st-place system of DCASE 2019 Task 4, which adopts weakly-supervised framework with an attention-based embedding-level pooling module and a semi-supervised learning approach named guided learning. This year, we incorporate multi-branch learning (MBL) into the original system to further improve its performance. MBL uses different branches with different pooling strategies (including instance-level and embedding-level strategies) and different pooling modules (including attention pooling, global max pooling or global average pooling modules), which share the same feature encoder of the model. Therefore, multiple branches pursuing different purposes and focusing on different characteristics of the data can help the feature encoder model the feature space better and avoid over-fitting. To better exploit the strongly-labeled synthetic data, inspired by multi-task learning, we also employ a sound event detection branch. To combine sound separation (SS) with sound event detection (SED), we fuse the results of SED systems with SS-SED systems which are trained using separated sound output by an SS system. The experimental results prove that MBL can improve the model performance and using SS has great potential to improve the performance of SED ensemble system.

Comments:	Accepted by DCASE2020 Workshop
Subjects:	Sound (cs.SD); Audio and Speech Processing (eess.AS)
Cite as:	arXiv:2007.10638 [cs.SD]
	(or arXiv:2007.10638v2 [cs.SD] for this version)
	https://doi.org/10.48550/arXiv.2007.10638

Submission history

From: Yuxin Huang [view email]
[v1] Tue, 21 Jul 2020 07:35:16 UTC (569 KB)
[v2] Mon, 2 Nov 2020 02:56:27 UTC (919 KB)

Computer Science > Sound

Title:Guided multi-branch learning systems for sound event detection with sound separation

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Sound

Title:Guided multi-branch learning systems for sound event detection with sound separation

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators