Pure and Spurious Critical Points: a Geometric Study of Linear Networks

Trager, Matthew; Kohn, Kathlén; Bruna, Joan

Computer Science > Machine Learning

arXiv:1910.01671 (cs)

[Submitted on 3 Oct 2019 (v1), last revised 3 Apr 2020 (this version, v2)]

Title:Pure and Spurious Critical Points: a Geometric Study of Linear Networks

Authors:Matthew Trager, Kathlén Kohn, Joan Bruna

View PDF

Abstract:The critical locus of the loss function of a neural network is determined by the geometry of the functional space and by the parameterization of this space by the network's weights. We introduce a natural distinction between pure critical points, which only depend on the functional space, and spurious critical points, which arise from the parameterization. We apply this perspective to revisit and extend the literature on the loss function of linear neural networks. For this type of network, the functional space is either the set of all linear maps from input to output space, or a determinantal variety, i.e., a set of linear maps with bounded rank. We use geometric properties of determinantal varieties to derive new results on the landscape of linear networks with different loss functions and different parameterizations. Our analysis clearly illustrates that the absence of "bad" local minima in the loss landscape of linear networks is due to two distinct phenomena that apply in different settings: it is true for arbitrary smooth convex losses in the case of architectures that can express all linear maps ("filling architectures") but it holds only for the quadratic loss when the functional space is a determinantal variety ("non-filling architectures"). Without any assumption on the architecture, smooth convex losses may lead to landscapes with many bad minima.

Subjects:	Machine Learning (cs.LG); Algebraic Geometry (math.AG); Machine Learning (stat.ML)
Cite as:	arXiv:1910.01671 [cs.LG]
	(or arXiv:1910.01671v2 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.1910.01671

Submission history

From: Matthew Trager [view email]
[v1] Thu, 3 Oct 2019 18:22:30 UTC (157 KB)
[v2] Fri, 3 Apr 2020 02:46:46 UTC (246 KB)

Computer Science > Machine Learning

Title:Pure and Spurious Critical Points: a Geometric Study of Linear Networks

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Pure and Spurious Critical Points: a Geometric Study of Linear Networks

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators