paper-with-me

홈 › Papers

BRICS: Bi-level feature Representation of Image CollectionS

2023-05-29 · Dingdong Yang, Yizhi Wang, Ali Mahdavi-Amiri, Hao Zhang

We present BRICS, a bi-level feature representation for image collections, which consists of a key code space on top of a feature grid space. Specifically, our representation is learned by an autoencoder to encode images into continuous key codes, which are used to retrieve features from groups of multi-resolution feature grids. Our key codes and feature grids are jointly trained continuously with well-defined gradient flows, leading to high usage rates of the feature grids and improved generative modeling compared to discrete Vector Quantization (VQ). Differently from existing continuous representations such as KL-regularized latent codes, our key codes are strictly bounded in scale and variance. Overall, feature encoding by BRICS is compact, efficient to train, and enables generative modeling over key codes using the diffusion model. Experimental results show that our method achieves comparable reconstruction results to VQ while having a smaller and more efficient decoder network (50% fewer GFlops). By applying the diffusion model over our key code space, we achieve state-of-the-art performance on image synthesis on the FFHQ and LSUN-Church (29% lower than LDM, 32% lower than StyleGAN2, 44% lower than Projected GAN on CLIP-FID) datasets.

📄 PDF Abstract BibTeX arXiv:2305.18601

Code (0)

등록된 구현이 없습니다.

Tasks

DecoderImage GenerationQuantization

Methods 이 논문이 사용한 방법론

Weight Demodulation 설명 없음
Path Length Regularization 설명 없음
R1 Regularization R_INLINE_MATH_1 Regularization is a regularization technique and gradient penalty for training [generative adversarial…
Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…
HuMan(Expedia)||How do I get a human at Expedia? How do I get a human at Expedia? How Do I Get a Human at Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Real-Time Help & Exclusive…
Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Using Analytic Scoring Rubrics in the Automatic Assessment of College-Level Summary Writing Tasks in L2

2017-11-01 · IJCNLP 2017 11 · Tamara Sladoljev-Agejev, Jan {\v{S}}najder

Assessing summaries is a demanding, yet useful task which provides valuable information on language competence, especially for second language learners. We consider automated scoring of college-level summary writing task…

Reading Comprehensionregression

One-Class Model for Fabric Defect Detection

2022-04-20 · Hao Zhou, Yixin Chen, David Troendle, Byunghyun Jang

An automated and accurate fabric defect inspection system is in high demand as a replacement for slow, inconsistent, error-prone, and expensive human operators in the textile industry. Previous efforts focused on certain…

Defect Detectionmodel

Fabric Surface Characterization: Assessment of Deep Learning-based Texture Representations Using a Challenging Dataset

2020-03-16 · Yuting Hu, Zhiling Long, Anirudha Sundaresan, Motaz Alfarraj 외

Tactile sensing or fabric hand plays a critical role in an individual's decision to buy a certain fabric from the range of available fabrics for a desired application. Therefore, textile and clothing manufacturers have l…

Material RecognitionObject RecognitionScene UnderstandingTexture Classification

Focus on the Positives: Self-Supervised Learning for Biodiversity Monitoring

2021-08-14 · ICCV 2021 10 · Omiros Pantazis, Gabriel Brostow, Kate Jones, Oisin Mac Aodha

We address the problem of learning self-supervised representations from unlabeled image collections. Unlike existing approaches that attempt to learn useful features by maximizing similarity between augmented versions of…

Self-Supervised LearningTransfer Learning

Learning Portrait Style Representations

2020-12-08 · Sadat Shaik, Bernadette Bucher, Nephele Agrafiotis, Stephen Phillips 외

Style analysis of artwork in computer vision predominantly focuses on achieving results in target image generation through optimizing understanding of low level style characteristics such as brush strokes. However, funda…

Image Generationzero-shot-classificationZero-Shot Learning