| J. Carreira, A. Zisserman | 2017 | Quo vadis, action recognition? A new model and the Kinetics dataset |
| J. Jacobsen, J. van Gemert, Z. Lou, A.W.M. Smeulders | 2016 | Structured receptive fields in CNNs |
| A. Karpathy, J. Johnson, L. Fei-Fei | 2015 | Visualizing and understanding recurrent networks |
| Q. Le, T. Mikolov | 2014 | Distributed representations of sentences and documents |
| Y. LeCun, Y. Bengio, G.E. Hinton | 2015 | Deep learning |
| C. Lin, S. Lucey | 2017 | Inverse compositional spatial transformer networks |
| J. Long, E. Shelhamer, T. Darrell | 2015 | Fully convolutional networks for semantic segmentation |
| G. Marcus | 2018 | Deep learning: a critical appraisal |
| P. Morerio, J. Cavazza, R. Volpi, R. Vidal, V. Murino | 2017 | Curriculum dropout |
| A. van den Oord, S. Dieleman, H. Zen, K. Simonyan, O. Vinyals, A. Graves, N. Kalchbrenner, A. Senior, K. Kavukcuoglu | 2016 | WaveNet: a generative model for raw audio |
| C.R. Qi, H. Su, K. Mo, L.J. Guibas | 2017 | PointNet: learning on point sets for 3D classification and segmentation |
| A.S. Razavian, H. Azizpour, J. Sullivan, S. Carlsson | 2014 | CNN features off-the-shelf: an astounding baseline for recognition |
| J. Redmon, S. Divvala, R. Girshick, A. Farhadi | 2016 | You only look once: unified, real-time object detection |
| S. Sabour, N. Frosst, G.E. Hinton | 2017 | Dynamic routing between capsules |
| A.M. Saxe, P.W. Koh, Z. Chen, M. Bhand, B. Suresh, A.Y. Ng | 2010 | On random weights and unsupervised feature learning |
| C. Szegedy, W. Zaremba, I. Sutskever, J. Bruna, D. Erhan, I. Goodfellow, R. Fergus | 2014 | Intriguing properties of neural networks |
| A. Veit, M. Wilber, S. Belongie | 2016 | Residual networks behave like ensembles of relatively shallow networks |
| K. Xu, J.L. Ba, R. Kiros, K. Cho, A. Courville, R. Salakhutdinov, R.S. Zemel, Y. Bengio | 2015 | Show, attend and tell: neural image caption generation with visual attention |
| M.D. Zeiler, R. Fergus | 2013 | Visualizing and understanding convolutional networks |
| R. Zhang, P. Isola, A.A. Efros | 2017 | Split-brain autoencoders: unsupervised learning by cross-channel prediction |
| J. Zhu, T. Park, P. Isola, A.A. Efros | 2017 | Unpaired image-to-image translation using cycle-consistent adversarial networks |