Weak-to-Strong Generalization Through the Data-Centric Lens

Shin, Changho; Cooper, John; Sala, Frederic

Computer Science > Machine Learning

arXiv:2412.03881 (cs)

[Submitted on 5 Dec 2024 (v1), last revised 4 Mar 2025 (this version, v2)]

Title:Weak-to-Strong Generalization Through the Data-Centric Lens

Authors:Changho Shin, John Cooper, Frederic Sala

View PDF HTML (experimental)

Abstract:The weak-to-strong generalization phenomenon is the driver for important machine learning applications including highly data-efficient learning and, most recently, performing superalignment. While decades of research have resulted in numerous algorithms that produce strong empirical performance, understanding what aspects of data enable weak-to-strong generalization has been understudied. We propose a simple data-centric mechanism that characterizes weak-to-strong generalization: the overlap density. Intuitively, generalization tracks the number of points that contain overlaps, i.e., both easy patterns (learnable by a weak model) and challenging patterns (only learnable by a stronger model), as with such points, weak predictions can be used to learn challenging patterns by stronger models. We provide a practical overlap detection algorithm to find such points in datasets and leverage them to learn, among multiple sources of data, which to query when seeking to maximize overlap density and thereby enhance weak-to-strong generalization. We present a theoretical result showing that the generalization benefit is a function of the overlap density and a regret bound for our data selection algorithm. Empirically, we validate the mechanism and the overlap detection algorithm on a wide array of settings.

Comments:	ICLR 2025
Subjects:	Machine Learning (cs.LG); Artificial Intelligence (cs.AI); Machine Learning (stat.ML)
Cite as:	arXiv:2412.03881 [cs.LG]
	(or arXiv:2412.03881v2 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2412.03881

Submission history

From: Changho Shin [view email]
[v1] Thu, 5 Dec 2024 05:29:19 UTC (2,936 KB)
[v2] Tue, 4 Mar 2025 04:28:19 UTC (2,937 KB)

Computer Science > Machine Learning

Title:Weak-to-Strong Generalization Through the Data-Centric Lens

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Weak-to-Strong Generalization Through the Data-Centric Lens

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators