Tradeoffs in Streaming Binary Classification under Limited Inspection Resources

Hassanzadeh, Parisa; Dervovic, Danial; Assefa, Samuel; Reddy, Prashant; Veloso, Manuela

Computer Science > Machine Learning

arXiv:2110.02403 (cs)

[Submitted on 5 Oct 2021 (v1), last revised 29 Oct 2021 (this version, v2)]

Title:Tradeoffs in Streaming Binary Classification under Limited Inspection Resources

Authors:Parisa Hassanzadeh, Danial Dervovic, Samuel Assefa, Prashant Reddy, Manuela Veloso

View PDF

Abstract:Institutions are increasingly relying on machine learning models to identify and alert on abnormal events, such as fraud, cyber attacks and system failures. These alerts often need to be manually investigated by specialists. Given the operational cost of manual inspections, the suspicious events are selected by alerting systems with carefully designed thresholds. In this paper, we consider an imbalanced binary classification problem, where events arrive sequentially and only a limited number of suspicious events can be inspected. We model the event arrivals as a non-homogeneous Poisson process, and compare various suspicious event selection methods including those based on static and adaptive thresholds. For each method, we analytically characterize the tradeoff between the minority-class detection rate and the inspection capacity as a function of the data class imbalance and the classifier confidence score densities. We implement the selection methods on a real public fraud detection dataset and compare the empirical results with analytical bounds. Finally, we investigate how class imbalance and the choice of classifier impact the tradeoff.

Comments:	To appear in Proceedings of the ACM International Conference on AI in Finance (ICAIF '21)
Subjects:	Machine Learning (cs.LG); Artificial Intelligence (cs.AI)
Cite as:	arXiv:2110.02403 [cs.LG]
	(or arXiv:2110.02403v2 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2110.02403

Submission history

From: Parisa Hassanzadeh [view email]
[v1] Tue, 5 Oct 2021 23:23:11 UTC (207 KB)
[v2] Fri, 29 Oct 2021 21:12:33 UTC (205 KB)

Computer Science > Machine Learning

Title:Tradeoffs in Streaming Binary Classification under Limited Inspection Resources

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Tradeoffs in Streaming Binary Classification under Limited Inspection Resources

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators