about
DecomposeMe: Simplifying ConvNets for End-To-End Learning (arxiv.org)
2 points by inferoscope on Oct 3, 2016 | hide | past | pdf | discuss on HN

In plain words: Instead of square filters, this network learns image features with thin one-way strips reused across layers, cutting memory and speed for small devices. On scene recognition it boosted relative top-1 accuracy by 7.7% while using 92% fewer parameters than a typical large network.

Abstract · DecomposeMe: Simplifying ConvNets for End-to-End Learning

Deep learning and convolutional neural networks (ConvNets) have been successfully applied to most relevant tasks in the computer vision community. However, these networks are computationally demanding and not suitable for embedded devices where memory and time consumption are relevant. In this paper, we propose DecomposeMe, a simple but effective technique to learn features using 1D convolutions. The proposed architecture enables both simplicity and filter sharing leading to increased learning capacity. A comprehensive set of large-scale experiments on ImageNet and Places2 demonstrates the ability of our method to improve performance while significantly reducing the number of parameters required. Notably, on Places2, we obtain an improvement in relative top-1 classification accuracy of 7.7\% with an architecture that requires 92% fewer parameters compared to VGG-B. The proposed network is also demonstrated to generalize to other tasks by converting existing networks.

Jose Alvarez, Lars Petersson
arXiv:1606.05426 · cs.CV · submitted Jun 17, 2016
abstract · pdf · html

add comment on HN