paper-with-me

홈 › Papers

The ImageNet Shuffle: Reorganized Pre-training for Video Event Detection

2016-02-23 · Pascal Mettes, Dennis C. Koelma, Cees G. M. Snoek

This paper strives for video event detection using a representation learned from deep convolutional neural networks. Different from the leading approaches, who all learn from the 1,000 classes defined in the ImageNet Large Scale Visual Recognition Challenge, we investigate how to leverage the complete ImageNet hierarchy for pre-training deep networks. To deal with the problems of over-specific classes and classes with few images, we introduce a bottom-up and top-down approach for reorganization of the ImageNet hierarchy based on all its 21,814 classes and more than 14 million images. Experiments on the TRECVID Multimedia Event Detection 2013 and 2015 datasets show that video representations derived from the layers of a deep neural network pre-trained with our reorganized hierarchy i) improves over standard pre-training, ii) is complementary among different reorganizations, iii) maintains the benefits of fusion with other modalities, and iv) leads to state-of-the-art event detection results. The reorganized hierarchies and their derived Caffe models are publicly available at http://tinyurl.com/imagenetshuffle.

📄 PDF Abstract BibTeX arXiv:1602.07119

Code (0)

등록된 구현이 없습니다.

Tasks

Event DetectionObject Recognition

Similar Papers 제목 키워드 기반

Stochastic Layer-Wise Shuffle: A Good Practice to Improve Vision Mamba Training

2024-08-30 · Zizheng Huang, Haoxing Chen, Jiaqi Li, Jun Lan 외

Recent Vision Mamba models not only have much lower complexity for processing higher resolution images and longer videos but also the competitive performance with Vision Transformers (ViTs). However, they are stuck into …

Image ClassificationMambaObject DetectionSemantic Segmentation

ShuffleNet: An Extremely Efficient Convolutional Neural Network for Mobile Devices

2017-07-04 · CVPR 2018 6 · Xiangyu Zhang, Xinyu Zhou, Mengxiao Lin, Jian Sun

We introduce an extremely computation-efficient CNN architecture named ShuffleNet, which is designed specially for mobile devices with very limited computing power (e.g., 10-150 MFLOPs). The new architecture utilizes two…

General ClassificationImage ClassificationObject DetectionPerson Re-Identification

Dynamic Shuffle: An Efficient Channel Mixture Method

2023-10-04 · Kaijun Gong, Zhuowen Yin, Yushu Li, Kailing Guo 외

The redundancy of Convolutional neural networks not only depends on weights but also depends on inputs. Shuffling is an efficient operation for mixing channel information but the shuffle order is usually pre-defined. To …

Binarizationimage-classificationImage Classification

ShuffleBlock: Shuffle to Regularize Deep Convolutional Neural Networks

2021-06-17 · Sudhakar Kumawat, Gagan Kanojia, Shanmuganathan Raman

Deep neural networks have enormous representational power which leads them to overfit on most datasets. Thus, regularizing them is important in order to reduce overfitting and enhance their generalization capabilities. R…

image-classificationImage ClassificationScheduling

Learning Efficient Video Representation with Video Shuffle Networks

2019-11-26 · Pingchuan Ma, Yao Zhou, Yu Lu, Wei zhang

3D CNN shows its strong ability in learning spatiotemporal representation in recent video recognition tasks. However, inflating 2D convolution to 3D inevitably introduces additional computational costs, making it cumbers…

Video Recognition