On the Accuracy of Influence Functions for Measuring Group Effects

Koh, Pang Wei; Ang, Kai-Siang; Teo, Hubert H. K.; Liang, Percy

Computer Science > Machine Learning

arXiv:1905.13289 (cs)

[Submitted on 30 May 2019 (v1), last revised 21 Nov 2019 (this version, v2)]

Title:On the Accuracy of Influence Functions for Measuring Group Effects

Authors:Pang Wei Koh, Kai-Siang Ang, Hubert H. K. Teo, Percy Liang

View PDF

Abstract:Influence functions estimate the effect of removing a training point on a model without the need to retrain. They are based on a first-order Taylor approximation that is guaranteed to be accurate for sufficiently small changes to the model, and so are commonly used to study the effect of individual points in large datasets. However, we often want to study the effects of large groups of training points, e.g., to diagnose batch effects or apportion credit between different data sources. Removing such large groups can result in significant changes to the model. Are influence functions still accurate in this setting? In this paper, we find that across many different types of groups and for a range of real-world datasets, the predicted effect (using influence functions) of a group correlates surprisingly well with its actual effect, even if the absolute and relative errors are large. Our theoretical analysis shows that such strong correlation arises only under certain settings and need not hold in general, indicating that real-world datasets have particular properties that allow the influence approximation to be accurate.

Subjects:	Machine Learning (cs.LG); Machine Learning (stat.ML)
Cite as:	arXiv:1905.13289 [cs.LG]
	(or arXiv:1905.13289v2 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.1905.13289

Submission history

From: Pang Wei Koh [view email]
[v1] Thu, 30 May 2019 20:24:17 UTC (2,056 KB)
[v2] Thu, 21 Nov 2019 06:49:54 UTC (2,065 KB)

Computer Science > Machine Learning

Title:On the Accuracy of Influence Functions for Measuring Group Effects

Submission history

Access Paper:

Current browse context:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:On the Accuracy of Influence Functions for Measuring Group Effects

Submission history

Access Paper:

Current browse context:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators