"Lossless" Compression of Deep Neural Networks: A High-dimensional Neural Tangent Kernel Approach

Gu, Lingyu; Du, Yongqi; Zhang, Yuan; Xie, Di; Pu, Shiliang; Qiu, Robert C.; Liao, Zhenyu

Statistics > Machine Learning

arXiv:2403.00258 (stat)

[Submitted on 1 Mar 2024]

Title:"Lossless" Compression of Deep Neural Networks: A High-dimensional Neural Tangent Kernel Approach

Authors:Lingyu Gu, Yongqi Du, Yuan Zhang, Di Xie, Shiliang Pu, Robert C. Qiu, Zhenyu Liao

View PDF

Abstract:Modern deep neural networks (DNNs) are extremely powerful; however, this comes at the price of increased depth and having more parameters per layer, making their training and inference more computationally challenging. In an attempt to address this key limitation, efforts have been devoted to the compression (e.g., sparsification and/or quantization) of these large-scale machine learning models, so that they can be deployed on low-power IoT devices. In this paper, building upon recent advances in neural tangent kernel (NTK) and random matrix theory (RMT), we provide a novel compression approach to wide and fully-connected \emph{deep} neural nets. Specifically, we demonstrate that in the high-dimensional regime where the number of data points $n$ and their dimension $p$ are both large, and under a Gaussian mixture model for the data, there exists \emph{asymptotic spectral equivalence} between the NTK matrices for a large family of DNN models. This theoretical result enables "lossless" compression of a given DNN to be performed, in the sense that the compressed network yields asymptotically the same NTK as the original (dense and unquantized) network, with its weights and activations taking values \emph{only} in $\{ 0, \pm 1 \}$ up to a scaling. Experiments on both synthetic and real-world data are conducted to support the advantages of the proposed compression scheme, with code available at \url{this https URL}.

Comments:	32 pages, 4 figures, and 2 tables. Fixing typos in Theorems 1 and 2 from NeurIPS 2022 proceeding (this https URL)
Subjects:	Machine Learning (stat.ML); Machine Learning (cs.LG)
Cite as:	arXiv:2403.00258 [stat.ML]
	(or arXiv:2403.00258v1 [stat.ML] for this version)
	https://doi.org/10.48550/arXiv.2403.00258

Submission history

From: Zhenyu Liao [view email]
[v1] Fri, 1 Mar 2024 03:46:28 UTC (158 KB)

Statistics > Machine Learning

Title:"Lossless" Compression of Deep Neural Networks: A High-dimensional Neural Tangent Kernel Approach

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Statistics > Machine Learning

Title:"Lossless" Compression of Deep Neural Networks: A High-dimensional Neural Tangent Kernel Approach

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators