资源论文V4D:4D CONVOLUTIONAL NEURAL NETWORKS FORV IDEO -LEVEL REPRESENTATIONS LEARNING

V4D:4D CONVOLUTIONAL NEURAL NETWORKS FORV IDEO -LEVEL REPRESENTATIONS LEARNING

2019-12-30 | |  72 |   53 |   0

Abstract

Most existing 3D CNN structures for video representation learning are clip-based methods, and do not consider video-level temporal evolution of spatio-temporal features. In this paper, we propose Video-level 4D Convolutional Neural Networks, namely V4D, to model the evolution of long-range spatio-temporal representation with 4D convolutions, as well as preserving 3D spatio-temporal representations with residual connections. We further introduce the training and inference methods for the proposed V4D. Extensive experiments are conducted on three video recognition benchmarks, where V4D achieves excellent results, surpassing recent 3D CNNs by a large margin.

上一篇:TOWARDS STABLE AND EFFICIENT TRAINING OFV ERIFIABLY ROBUST NEURAL NETWORKS

下一篇:PROVABLE FILTER PRUNING FOR EFFICIENT NEURAL N ETWORKS

用户评价
全部评价

热门资源

  • Learning to Predi...

    Much of model-based reinforcement learning invo...

  • Stratified Strate...

    In this paper we introduce Stratified Strategy ...

  • The Variational S...

    Unlike traditional images which do not offer in...

  • A Mathematical Mo...

    Direct democracy, where each voter casts one vo...

  • Rating-Boosted La...

    The performance of a recommendation system reli...