Feature Noise Induces Loss Discrepancy Across Groups

Khani, Fereshte; Liang, Percy

Computer Science > Machine Learning

arXiv:1911.09876 (cs)

[Submitted on 22 Nov 2019 (v1), last revised 6 Nov 2020 (this version, v2)]

Title:Feature Noise Induces Loss Discrepancy Across Groups

Authors:Fereshte Khani, Percy Liang

View PDF

Abstract:The performance of standard learning procedures has been observed to differ widely across groups. Recent studies usually attribute this loss discrepancy to an information deficiency for one group (e.g., one group has less data). In this work, we point to a more subtle source of loss discrepancy---feature noise. Our main result is that even when there is no information deficiency specific to one group (e.g., both groups have infinite data), adding the same amount of feature noise to all individuals leads to loss discrepancy. For linear regression, we thoroughly characterize the effect of feature noise on loss discrepancy in terms of the amount of noise, the difference between moments of the two groups, and whether group information is used or not. We then show this loss discrepancy does not vanish immediately if a shift in distribution causes the groups to have similar moments. On three real-world datasets, we show feature noise increases the loss discrepancy if groups have different distributions, while it does not affect the loss discrepancy on datasets where groups have similar distributions.

Comments:	ICML 2020
Subjects:	Machine Learning (cs.LG); Machine Learning (stat.ML)
Cite as:	arXiv:1911.09876 [cs.LG]
	(or arXiv:1911.09876v2 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.1911.09876

Submission history

From: Fereshte Khani [view email]
[v1] Fri, 22 Nov 2019 06:36:23 UTC (1,962 KB)
[v2] Fri, 6 Nov 2020 01:51:22 UTC (5,470 KB)

Full-text links:

Access Paper:

view license

Current browse context:

cs.LG

< prev | next >

new | recent | 2019-11

Change to browse by:

cs
stat
stat.ML

References & Citations

DBLP - CS Bibliography

listing | bibtex

Fereshte Khani
Percy Liang

export BibTeX citation

Computer Science > Machine Learning

Title:Feature Noise Induces Loss Discrepancy Across Groups

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Feature Noise Induces Loss Discrepancy Across Groups

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators