Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Efficient Sparse-Winograd Convolutional Neural Networks

Feb 18, 2018

Xingyu Liu, Jeff Pool, Song Han, William J. Dally

Figure 1 for Efficient Sparse-Winograd Convolutional Neural Networks

Figure 2 for Efficient Sparse-Winograd Convolutional Neural Networks

Figure 3 for Efficient Sparse-Winograd Convolutional Neural Networks

Figure 4 for Efficient Sparse-Winograd Convolutional Neural Networks

Share this with someone who'll enjoy it:

Abstract:Convolutional Neural Networks (CNNs) are computationally intensive, which limits their application on mobile devices. Their energy is dominated by the number of multiplies needed to perform the convolutions. Winograd's minimal filtering algorithm (Lavin, 2015) and network pruning (Han et al., 2015) can reduce the operation count, but these two methods cannot be directly combined $-$ applying the Winograd transform fills in the sparsity in both the weights and the activations. We propose two modifications to Winograd-based CNNs to enable these methods to exploit sparsity. First, we move the ReLU operation into the Winograd domain to increase the sparsity of the transformed activations. Second, we prune the weights in the Winograd domain to exploit static weight sparsity. For models on CIFAR-10, CIFAR-100 and ImageNet datasets, our method reduces the number of multiplications by $10.4\times$, $6.8\times$ and $10.8\times$ respectively with loss of accuracy less than $0.1\%$, outperforming previous baselines by $2.0\times$-$3.0\times$. We also show that moving ReLU to the Winograd domain allows more aggressive pruning.

* Published as a conference paper at ICLR 2018

View paper on

OpenReview

Share this with someone who'll enjoy it:

Title:Efficient Sparse-Winograd Convolutional Neural Networks

Paper and Code