about
QuickNet: Maximizing Efficiency and Efficacy in Deep Architectures (arxiv.org)
2 points by fuwuxopape on Jan 10, 2017 | hide | past | pdf | discuss on HN

In plain words: QuickNet speeds up a standard image network by splitting its filters into cheaper pieces and letting each unit pass a small signal when off. It reaches 95.7 percent accuracy on a common image test while running faster and using less memory than other quick networks.

Abstract

We present QuickNet, a fast and accurate network architecture that is both faster and significantly more accurate than other fast deep architectures like SqueezeNet. Furthermore, it uses less parameters than previous networks, making it more memory efficient. We do this by making two major modifications to the reference Darknet model (Redmon et al, 2015): 1) The use of depthwise separable convolutions and 2) The use of parametric rectified linear units. We make the observation that parametric rectified linear units are computationally equivalent to leaky rectified linear units at test time and the observation that separable convolutions can be interpreted as a compressed Inception network (Chollet, 2016). Using these observations, we derive a network architecture, which we call QuickNet, that is both faster and more accurate than previous models. Our architecture provides at least four major advantages: (1) A smaller model size, which is more tenable on memory constrained systems; (2) A significantly faster network which is more tenable on computationally constrained systems; (3) A high accuracy of 95.7 percent on the CIFAR-10 Dataset which outperforms all but one result published so far, although we note that our works are orthogonal approaches and can be combined (4) Orthogonality to previous model compression approaches allowing for further speed gains to be realized.

Tapabrata Ghosh
arXiv:1701.02291 · cs.LG, stat.ML · submitted Jan 9, 2017 · updated Jan 12, 2017
abstract · pdf · Updated once

add comment on HN