[–] ekelsen 6y ago ↗ TLDR; Make 1x1 convolutions sparse, write fast Sparse Matrix Multiplication kernels, get a nearly 2x speedup with smaller models.
1 comment
[ 2.9 ms ] story [ 20.6 ms ] thread