Improving Label Quality by Jointly Modeling Items and Annotators

Weerasooriya, Tharindu Cyril; Ororbia, Alexander G.; Homan, Christopher M.

Computer Science > Artificial Intelligence

arXiv:2106.10600 (cs)

[Submitted on 20 Jun 2021]

Title:Improving Label Quality by Jointly Modeling Items and Annotators

Authors:Tharindu Cyril Weerasooriya, Alexander G. Ororbia, Christopher M. Homan

View PDF

Abstract:We propose a fully Bayesian framework for learning ground truth labels from noisy annotators.
Our framework ensures scalability by factoring a generative, Bayesian soft clustering model over label distributions into the classic David and Skene joint annotator-data model. Earlier research along these lines has neither fully incorporated label distributions nor explored clustering by annotators only or data only. Our framework incorporates all of these properties as:
(1) a graphical model designed to provide better ground truth estimates of annotator responses as input to \emph{any} black box supervised learning algorithm, and
(2) a standalone neural model whose internal structure captures many of the properties of the graphical model.
We conduct supervised learning experiments using both models and compare them to the performance of one baseline and a state-of-the-art model.

Subjects:	Artificial Intelligence (cs.AI); Social and Information Networks (cs.SI)
Cite as:	arXiv:2106.10600 [cs.AI]
	(or arXiv:2106.10600v1 [cs.AI] for this version)
	https://doi.org/10.48550/arXiv.2106.10600

Submission history

From: Tharindu Cyril Weerasooriya [view email]
[v1] Sun, 20 Jun 2021 02:15:20 UTC (1,731 KB)

Full-text links:

Access Paper:

view license

Current browse context:

cs.AI

< prev | next >

new | recent | 2021-06

Change to browse by:

cs
cs.SI

References & Citations

DBLP - CS Bibliography

listing | bibtex

Alexander G. Ororbia II
Christopher M. Homan

export BibTeX citation

Computer Science > Artificial Intelligence

Title:Improving Label Quality by Jointly Modeling Items and Annotators

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Artificial Intelligence

Title:Improving Label Quality by Jointly Modeling Items and Annotators

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators