paper-with-me

Papers

Self-supervised Photographic Image Layout Representation Learning

2024-03-06 · Zhaoran Zhao, Peng Lu, Xujun Peng, Wenhao Guo

In the domain of image layout representation learning, the critical process of translating image layouts into succinct vector forms is increasingly significant across diverse applications, such as image retrieval, manipulation, and generation. Most approaches in this area heavily rely on costly labeled datasets and notably lack in adapting their modeling and learning methods to the specific nuances of photographic image layouts. This shortfall makes the learning process for photographic image layouts suboptimal. In our research, we directly address these challenges. We innovate by defining basic layout primitives that encapsulate various levels of layout information and by mapping these, along with their interconnections, onto a heterogeneous graph structure. This graph is meticulously engineered to capture the intricate layout information within the pixel domain explicitly. Advancing further, we introduce novel pretext tasks coupled with customized loss functions, strategically designed for effective self-supervised learning of these layout graphs. Building on this foundation, we develop an autoencoder-based network architecture skilled in compressing these heterogeneous layout graphs into precise, dimensionally-reduced layout representations. Additionally, we introduce the LODB dataset, which features a broader range of layout categories and richer semantics, serving as a comprehensive benchmark for evaluating the effectiveness of layout representation learning methods. Our extensive experimentation on this dataset demonstrates the superior performance of our approach in the realm of photographic image layout representation learning.

📄 PDF Abstract BibTeX arXiv:2403.03740

Code (1)

cv-xueba/image-layout-learning 공식 구현

Tasks

Image RetrievalRepresentation LearningSelf-Supervised Learning

Similar Papers 제목 키워드 기반

Photographic Image Synthesis with Cascaded Refinement Networks

2017-07-28 · ICCV 2017 10 · Qifeng Chen, Vladlen Koltun

We present an approach to synthesizing photographic images conditioned on semantic layouts. Given a semantic label map, our approach produces an image with photographic appearance that conforms to the input layout. The a…

Image GenerationImage-to-Image Translation

Semi-parametric Image Synthesis

2018-04-29 · CVPR 2018 6 · Xiaojuan Qi, Qifeng Chen, Jiaya Jia, Vladlen Koltun

We present a semi-parametric approach to photographic image synthesis from semantic layouts. The approach combines the complementary strengths of parametric and nonparametric techniques. The nonparametric component is a …

Image GenerationImage-to-Image TranslationSemantic Segmentation

Semantically Stable Image Composition Analysis via Saliency and Gradient Vector Flow Fusion

2026-04-14 · Armin Dadras, Robert Sablatnig, Franziska Proksa, Markus Seidl arxiv

The reliable computational assessment of photographic composition requires features that are discriminative of spatial layout yet robust to semantic content. This paper proposes a low-level representation grounded in the…

Self-supervised 360$^{\circ}$ Room Layout Estimation

2022-03-30 · Hao-Wen Ting, Cheng Sun, Hwann-Tzong Chen

We present the first self-supervised method to train panoramic room layout estimation models without any labeled data. Unlike per-pixel dense depth that provides abundant correspondence constraints, layout representation…

Active LearningRoom Layout Estimation

CAiD: Context-Aware Instance Discrimination for Self-supervised Learning in Medical Imaging

2022-04-15 · Mohammad Reza Hosseinzadeh Taher, Fatemeh Haghighi, Michael B. Gotway, Jianming Liang

Recently, self-supervised instance discrimination methods have achieved significant success in learning visual representations from unlabeled photographic images. However, given the marked differences between photographi…

AnatomySelf-Supervised Learning