paper-with-me

Papers

Fusing Deep Convolutional Networks for Large Scale Visual Concept Classification

2016-08-05 · Hilal Ergun, Mustafa Sert

Deep learning architectures are showing great promise in various computer vision domains including image classification, object detection, event detection and action recognition. In this study, we investigate various aspects of convolutional neural networks (CNNs) from the big data perspective. We analyze recent studies and different network architectures both in terms of running time and accuracy. We present extensive empirical information along with best practices for big data practitioners. Using these best practices we propose efficient fusion mechanisms both for single and multiple network models. We present state-of-the art results on benchmark datasets while keeping computational costs at a lower level. Another contribution of our paper is that these state-of-the-art results can be reached without using extensive data augmentation techniques.

📄 PDF Abstract BibTeX arXiv:1608.01866

Code (0)

등록된 구현이 없습니다.

Tasks

Action RecognitionClassificationData AugmentationEvent DetectionGeneral Classificationimage-classificationImage Classificationobject-detectionObject DetectionTemporal Action Localization

Similar Papers 제목 키워드 기반

Multi-Modal Multi-Scale Deep Learning for Large-Scale Image Annotation

2017-09-05 · Yulei Niu, Zhiwu Lu, Ji-Rong Wen, Tao Xiang 외

Image annotation aims to annotate a given image with a variable number of class labels corresponding to diverse visual concepts. In this paper, we address two main issues in large-scale image annotation: 1) how to learn …

Deep Learning

Boosting the accuracy of multi-spectral image pan-sharpening by learning a deep residual network

2017-05-22 · Yancong Wei, Qiangqiang Yuan, Huanfeng Shen, Liangpei Zhang

In the field of fusing multi-spectral and panchromatic images (Pan-sharpening), the impressive effectiveness of deep neural networks has been recently employed to overcome the drawbacks of traditional linear models and b…

Non-confusing Generation of Customized Concepts in Diffusion Models

2024-05-11 · Wang Lin, Jingyuan Chen, Jiaxin Shi, Yichen Zhu 외

We tackle the common challenge of inter-concept visual confusion in compositional concept generation using text-guided diffusion models (TGDMs). It becomes even more pronounced in the generation of customized concepts, d…

Regional Multi-scale Approach for Visually Pleasing Explanations of Deep Neural Networks

2018-07-31 · Dasom Seo, Kanghan Oh, Il-Seok Oh

Recently, many methods to interpret and visualize deep neural network predictions have been proposed and significant progress has been made. However, a more class-discriminative and visually pleasing explanation is requi…

Feature Importance

How intelligent are convolutional neural networks?

2017-09-18 · Zhennan Yan, Xiang Sean Zhou

Motivated by the Gestalt pattern theory, and the Winograd Challenge for language understanding, we design synthetic experiments to investigate a deep learning algorithm's ability to infer simple (at least for human) visu…