Skip to main content

Showing 1–3 of 3 results for author: Lau, E

Searching in archive stat. Search in all archives.
.
  1. arXiv:2402.03698   

    cs.LG stat.ML

    Estimating the Local Learning Coefficient at Scale

    Authors: Zach Furman, Edmund Lau

    Abstract: The \textit{local learning coefficient} (LLC) is a principled way of quantifying model complexity, originally derived in the context of Bayesian statistics using singular learning theory (SLT). Several methods are known for numerically estimating the local learning coefficient, but so far these methods have not been extended to the scale of modern deep learning architectures or data sets. Using a… ▽ More

    Submitted 30 September, 2024; v1 submitted 5 February, 2024; originally announced February 2024.

    Comments: This paper has been expanded and merged with arXiv:2308.12108 to form a more comprehensive study. Please refer to the latest version of that preprint for the most up-to-date manuscript

    MSC Class: 68T07; 14B05; 62F15

  2. arXiv:2308.12108  [pdf, other

    stat.ML cs.AI cs.LG

    The Local Learning Coefficient: A Singularity-Aware Complexity Measure

    Authors: Edmund Lau, Zach Furman, George Wang, Daniel Murfet, Susan Wei

    Abstract: The Local Learning Coefficient (LLC) is introduced as a novel complexity measure for deep neural networks (DNNs). Recognizing the limitations of traditional complexity measures, the LLC leverages Singular Learning Theory (SLT), which has long recognized the significance of singularities in the loss landscape geometry. This paper provides an extensive exploration of the LLC's theoretical underpinni… ▽ More

    Submitted 30 September, 2024; v1 submitted 23 August, 2023; originally announced August 2023.

    Comments: This version contains new empirical results and merged content from a related paper (arXiv:2402.03698) to provide a more comprehensive study

    MSC Class: 62F15; 68T07; 14B05

  3. arXiv:2302.06035  [pdf, other

    stat.ML cs.AI cs.LG

    Variational Bayesian Neural Networks via Resolution of Singularities

    Authors: Susan Wei, Edmund Lau

    Abstract: In this work, we advocate for the importance of singular learning theory (SLT) as it pertains to the theory and practice of variational inference in Bayesian neural networks (BNNs). To begin, using SLT, we lay to rest some of the confusion surrounding discrepancies between downstream predictive performance measured via e.g., the test log predictive density, and the variational objective. Next, we… ▽ More

    Submitted 12 February, 2023; originally announced February 2023.

    Comments: 32 pages, 13 figures

    MSC Class: 62F15 (Primary); 68T07 (Secondary); 68T05