Online Importance Sampling for Stochastic Gradient Optimization

Salaün, Corentin; Huang, Xingchang; Georgiev, Iliyan; Mitra, Niloy J.; Singh, Gurprit

Computer Science > Machine Learning

arXiv:2311.14468 (cs)

[Submitted on 24 Nov 2023 (v1), last revised 28 Jan 2025 (this version, v3)]

Title:Online Importance Sampling for Stochastic Gradient Optimization

Authors:Corentin Salaün, Xingchang Huang, Iliyan Georgiev, Niloy J. Mitra, Gurprit Singh

View PDF HTML (experimental)

Abstract:Machine learning optimization often depends on stochastic gradient descent, where the precision of gradient estimation is vital for model performance. Gradients are calculated from mini-batches formed by uniformly selecting data samples from the training dataset. However, not all data samples contribute equally to gradient estimation. To address this, various importance sampling strategies have been developed to prioritize more significant samples. Despite these advancements, all current importance sampling methods encounter challenges related to computational efficiency and seamless integration into practical machine learning pipelines. In this work, we propose a practical algorithm that efficiently computes data importance on-the-fly during training, eliminating the need for dataset preprocessing. We also introduce a novel metric based on the derivative of the loss w.r.t. the network output, designed for mini-batch importance sampling. Our metric prioritizes influential data points, thereby enhancing gradient estimation accuracy. We demonstrate the effectiveness of our approach across various applications. We first perform classification and regression tasks to demonstrate improvements in accuracy. Then, we show how our approach can also be used for online data pruning by identifying and discarding data samples that contribute minimally towards the training loss. This significantly reduce training time with negligible loss in the accuracy of the model.

Comments:	17 pages, 7 figures
Subjects:	Machine Learning (cs.LG)
Cite as:	arXiv:2311.14468 [cs.LG]
	(or arXiv:2311.14468v3 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2311.14468

Submission history

From: Corentin Salaun [view email]
[v1] Fri, 24 Nov 2023 13:21:35 UTC (4,227 KB)
[v2] Mon, 27 Nov 2023 08:04:04 UTC (4,227 KB)
[v3] Tue, 28 Jan 2025 09:29:21 UTC (22,711 KB)

Computer Science > Machine Learning

Title:Online Importance Sampling for Stochastic Gradient Optimization

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Online Importance Sampling for Stochastic Gradient Optimization

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators