Building effective deep neural network architectures one feature at a time

Mundt, Martin; Weis, Tobias; Konda, Kishore; Ramesh, Visvanathan

Computer Science > Computer Vision and Pattern Recognition

arXiv:1705.06778v1 (cs)

[Submitted on 18 May 2017 (this version), latest version 19 Oct 2017 (v2)]

Title:Building effective deep neural network architectures one feature at a time

Authors:Martin Mundt, Tobias Weis, Kishore Konda, Visvanathan Ramesh

View PDF

Abstract:Successful training of convolutional neural networks is often associated with the training of sufficiently deep architectures composed of high amounts of features while relying on a variety of regularization and pruning techniques to converge to less redundant states. We introduce an easy to compute metric, based on feature time evolution, to evaluate feature importance during training and demonstrate its potency in determining a networks effective capacity. In consequence we propose a novel algorithm to evolve fixed-depth architectures starting from just a single feature per layer to attain effective representational capacities needed for a specific task by greedily adding feature by feature. We revisit popular CNN architectures and demonstrate how evolved architectures not only converge to similar topologies that benefit from less parameters or improved accuracy, but furthermore exhibit systematic correspondence in representational complexity with the specified task. In contrast to conventional design patterns that typically have a monotonic increase in the amount of features with increased depth, we observe that CNNs perform better when there is a peak in learnable parameters in intermediate, with falloffs to earlier and later layers.

Subjects:	Computer Vision and Pattern Recognition (cs.CV); Neural and Evolutionary Computing (cs.NE)
Cite as:	arXiv:1705.06778 [cs.CV]
	(or arXiv:1705.06778v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.1705.06778

Submission history

From: Martin Mundt [view email]
[v1] Thu, 18 May 2017 19:40:37 UTC (1,431 KB)
[v2] Thu, 19 Oct 2017 21:59:52 UTC (307 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Building effective deep neural network architectures one feature at a time

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Building effective deep neural network architectures one feature at a time

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators