paper-with-me

홈 › Papers

Channel-Recurrent Autoencoding for Image Modeling

2017-06-12 · Wenling Shang, Kihyuk Sohn, Yuandong Tian

Despite recent successes in synthesizing faces and bedrooms, existing generative models struggle to capture more complex image types, potentially due to the oversimplification of their latent space constructions. To tackle this issue, building on Variational Autoencoders (VAEs), we integrate recurrent connections across channels to both inference and generation steps, allowing the high-level features to be captured in global-to-local, coarse-to-fine manners. Combined with adversarial loss, our channel-recurrent VAE-GAN (crVAE-GAN) outperforms VAE-GAN in generating a diverse spectrum of high resolution images while maintaining the same level of computational efficacy. Our model produces interpretable and expressive latent representations to benefit downstream tasks such as image completion. Moreover, we propose two novel regularizations, namely the KL objective weighting scheme over time steps and mutual information maximization between transformed latent variables and the outputs, to enhance the training.

📄 PDF Abstract BibTeX arXiv:1706.03729

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

End-to-End Optimized Transmission over Dispersive Intensity-Modulated Channels Using Bidirectional Recurrent Neural Networks

2019-01-24 · Boris Karanov, Domaniç Lavery, Polina Bayvel, Laurent Schmalen

We propose an autoencoding sequence-based transceiver for communication over dispersive channels with intensity modulation and direct detection (IM/DD), designed as a bidirectional deep recurrent neural network (BRNN). T…

Decoding 5G-NR Communications via Deep Learning

2020-07-15 · Pol Henarejos, Miguel Ángel Vázquez

Upcoming modern communications are based on 5G specifications and aim at providing solutions for novel vertical industries. One of the major changes of the physical layer is the use of Low-Density Parity-Check (LDPC) cod…

Deep Learning

Character-level Supervision for Low-resource POS Tagging

2018-07-01 · WS 2018 7 · Katharina Kann, Johannes Bjerva, Isabelle Augenstein, Barbara Plank 외

Neural part-of-speech (POS) taggers are known to not perform well with little training data. As a step towards overcoming this problem, we present an architecture for learning more robust neural POS taggers by jointly tr…

Feature EngineeringLEMMALemmatizationMulti-Task Learning+2

Bridging Autoencoders and Dynamic Mode Decomposition for Reduced-order Modeling and Control of PDEs

2024-09-09 · Priyabrata Saha, Saibal Mukhopadhyay

Modeling and controlling complex spatiotemporal dynamical systems driven by partial differential equations (PDEs) often necessitate dimensionality reduction techniques to construct lower-order models for computational ef…

Computational EfficiencyDimensionality Reduction

RP-CATE: Recurrent Perceptron-based Channel Attention Transformer Encoder for Industrial Hybrid Modeling

2025-12-22 · Haoran Yang, Yinan Zhang, Wenjie Zhang, Dongxia Wang 외 arxiv

Nowadays, industrial hybrid modeling which integrates both mechanistic modeling and machine learning-based modeling techniques has attracted increasing interest from scholars due to its high accuracy, low computational c…