paper-with-me

Papers

Semantic-Rearrangement-Based Multi-Level Alignment for Domain Generalized Segmentation

2024-04-21 · Guanlong Jiao, Chenyangguang Zhang, Haonan Yin, Yu Mo, Biqing Huang, Hui Pan, Yi Luo, Jingxian Liu

Domain generalized semantic segmentation is an essential computer vision task, for which models only leverage source data to learn the capability of generalized semantic segmentation towards the unseen target domains. Previous works typically address this challenge by global style randomization or feature regularization. In this paper, we argue that given the observation that different local semantic regions perform different visual characteristics from the source domain to the target domain, methods focusing on global operations are hard to capture such regional discrepancies, thus failing to construct domain-invariant representations with the consistency from local to global level. Therefore, we propose the Semantic-Rearrangement-based Multi-Level Alignment (SRMA) to overcome this problem. SRMA first incorporates a Semantic Rearrangement Module (SRM), which conducts semantic region randomization to enhance the diversity of the source domain sufficiently. A Multi-Level Alignment module (MLA) is subsequently proposed with the help of such diversity to establish the global-regional-local consistent domain-invariant representations. By aligning features across randomized samples with domain-neutral knowledge at multiple levels, SRMA provides a more robust way to handle the source-target domain gap. Extensive experiments demonstrate the superiority of SRMA over the current state-of-the-art works on various benchmarks.

📄 PDF Abstract BibTeX arXiv:2404.13701

Code (0)

등록된 구현이 없습니다.

Tasks

DiversitySemantic Segmentation

Similar Papers 제목 키워드 기반

PARSEC: Preference Adaptation for Robotic Object Rearrangement from Scene Context

2025-05-16 · Kartik Ramachandruni, Sonia Chernova

Object rearrangement is a key task for household robots requiring personalization without explicit instructions, meaningful object placement in environments occupied with objects, and generalization to unseen objects and…

ObjectObject Rearrangement

HGAN: Hierarchical Graph Alignment Network for Image-Text Retrieval

2022-12-16 · Jie Guo, Meiting Wang, Yan Zhou, Bin Song 외

Image-text retrieval (ITR) is a challenging task in the field of multimodal information processing due to the semantic gap between different modalities. In recent years, researchers have made great progress in exploring …

Image-text RetrievalRetrievalSentenceText Retrieval

A Simple Approach for Visual Rearrangement: 3D Mapping and Semantic Search

2022-06-21 · Brandon Trabucco, Gunnar Sigurdsson, Robinson Piramuthu, Gaurav S. Sukhatme 외

Physically rearranging objects is an important capability for embodied agents. Visual room rearrangement evaluates an agent's ability to rearrange objects in a room to a desired goal based solely on visual input. We prop…

Semantic Segmentation

Q-Align: Alleviating Attention Leakage in Zero-Shot Appearance Transfer via Query-Query Alignment

2025-08-27 · Namu Kim, Wonbin Kweon, Minsoo Kim, Hwanjo Yu arxiv

We observe that zero-shot appearance transfer with large-scale image generation models faces a significant challenge: Attention Leakage. This challenge arises when the semantic mapping between two images is captured by t…

Image Generation

Humanoid Hanoi: Investigating Shared Whole-Body Control for Skill-Based Box Rearrangement

2026-02-14 · Minku Kim, Kuan-Chia Chen, Aayam Shrestha, Li Fuxin 외 arxiv

We investigate a skill-based framework for humanoid box rearrangement that enables long-horizon execution by sequencing reusable skills at the task level. In our architecture, all skills execute through a shared, task-ag…