paper-with-me

Papers

Plug-in, Trainable Gate for Streamlining Arbitrary Neural Networks

2019-04-24 · Jaedeok Kim, Chiyoun Park, Hyun-Joo Jung, Yoonsuck Choe

Architecture optimization, which is a technique for finding an efficient neural network that meets certain requirements, generally reduces to a set of multiple-choice selection problems among alternative sub-structures or parameters. The discrete nature of the selection problem, however, makes this optimization difficult. To tackle this problem we introduce a novel concept of a trainable gate function. The trainable gate function, which confers a differentiable property to discretevalued variables, allows us to directly optimize loss functions that include non-differentiable discrete values such as 0-1 selection. The proposed trainable gate can be applied to pruning. Pruning can be carried out simply by appending the proposed trainable gate functions to each intermediate output tensor followed by fine-tuning the overall model, using any gradient-based training methods. So the proposed method can jointly optimize the selection of the pruned channels while fine-tuning the weights of the pruned model at the same time. Our experimental results demonstrate that the proposed method efficiently optimizes arbitrary neural networks in various tasks such as image classification, style transfer, optical flow estimation, and neural machine translation.

📄 PDF Abstract BibTeX arXiv:1904.10921

Code (0)

등록된 구현이 없습니다.

Tasks

Efficient Neural Networkimage-classificationImage ClassificationMachine TranslationMultiple-choiceOptical Flow EstimationStyle TransferTranslation

Methods 이 논문이 사용한 방법론

Pruning 설명 없음

Similar Papers 제목 키워드 기반

Towards Run Time Estimation of the Gaussian Chemistry Code for SEAGrid Science Gateway

2019-06-07 · Angel Beltre, Shehtab Zaman, Kenneth Chiu, Sudhakar Pamidighantam 외

Accurate estimation of the run time of computational codes has a number of significant advantages for scientific computing. It is required information for optimal resource allocation, improving turnaround times and utili…

Computational chemistry

Plug-In RIS: A Novel Approach to Fully Passive Reconfigurable Intelligent Surfaces

2023-11-16 · Mahmoud Raeisi, Ibrahim Yildirim, Mehmet Cagri Ilter, Majid Gerami 외

This paper presents a promising design concept for reconfigurable intelligent surfaces (RISs), named plug-in RIS, wherein the RIS is plugged into an appropriate position in the environment, adjusted once according to the…

Zero-Observation User Reactivation with Gap-Driven Dimensional Gating

2026-07-22 · Jiandong Ding, Tianying Liu, Fuyuan Liu, Huijie Qin 외 arxiv

Sequential recommendation (SR) models capture continuously observed behavior, but a returning user may have no interactions for months or years. We define this setting as Zero-Observation Reactivation: the user has a pre…

Sequential Recommendation

Deep Plug-and-Play Super-Resolution for Arbitrary Blur Kernels

2019-03-29 · CVPR 2019 6 · Kai Zhang, WangMeng Zuo, Lei Zhang

While deep neural networks (DNN) based single image super-resolution (SISR) methods are rapidly gaining popularity, they are mainly designed for the widely-used bicubic degradation, and there still remains the fundamenta…

DeblurringImage RestorationImage Super-ResolutionSuper-Resolution

Learning A Single Network for Scale-Arbitrary Super-Resolution

2020-04-08 · ICCV 2021 10 · Longguang Wang, Yingqian Wang, Zaiping Lin, Jungang Yang 외

Recently, the performance of single image super-resolution (SR) has been significantly improved with powerful networks. However, these networks are developed for image SR with a single specific integer scale (e.g., x2;x3…

Image Super-ResolutionSuper-ResolutionTransfer Learning