paper-with-me

홈 › Papers

Progressive Domain-Independent Feature Decomposition Network for Zero-Shot Sketch-Based Image Retrieval

2020-03-22 · Xinxun Xu, Muli Yang, Yanhua Yang, Hao Wang

Zero-shot sketch-based image retrieval (ZS-SBIR) is a specific cross-modal retrieval task for searching natural images given free-hand sketches under the zero-shot scenario. Most existing methods solve this problem by simultaneously projecting visual features and semantic supervision into a low-dimensional common space for efficient retrieval. However, such low-dimensional projection destroys the completeness of semantic knowledge in original semantic space, so that it is unable to transfer useful knowledge well when learning semantic from different modalities. Moreover, the domain information and semantic information are entangled in visual features, which is not conducive for cross-modal matching since it will hinder the reduction of domain gap between sketch and image. In this paper, we propose a Progressive Domain-independent Feature Decomposition (PDFD) network for ZS-SBIR. Specifically, with the supervision of original semantic knowledge, PDFD decomposes visual features into domain features and semantic ones, and then the semantic features are projected into common space as retrieval features for ZS-SBIR. The progressive projection strategy maintains strong semantic supervision. Besides, to guarantee the retrieval features to capture clean and complete semantic information, the cross-reconstruction loss is introduced to encourage that any combinations of retrieval features and domain features can reconstruct the visual features. Extensive experiments demonstrate the superiority of our PDFD over state-of-the-art competitors.

📄 PDF Abstract BibTeX arXiv:2003.09869

Code (0)

등록된 구현이 없습니다.

Tasks

Cross-Modal RetrievalImage RetrievalRetrievalSketch-Based Image Retrieval

Similar Papers 제목 키워드 기반

Independent Feature Decomposition and Instance Alignment for Unsupervised Domain Adaptation

2023-04-19 · IJCAI 2023 4 · Qichen He, Siying Xiao, Mao Ye, Xiatian Zhu 외

Existing Unsupervised Domain Adaptation (UDA) methods typically attempt to perform knowledge transfer in a domain-invariant space explicitly or implicitly. In practice, however, the obtained features are often mixed with…

Domain AdaptationTransfer LearningUnsupervised Domain Adaptation

Deformable Model-Driven Neural Rendering for High-Fidelity 3D Reconstruction of Human Heads Under Low-View Settings

2023-03-24 · ICCV 2023 1 · Baixin Xu, Jiarui Zhang, Kwan-Yee Lin, Chen Qian 외

Reconstructing 3D human heads in low-view settings presents technical challenges, mainly due to the pronounced risk of overfitting with limited views and high-frequency signals. To address this, we propose geometry decom…

3D ReconstructionNeural RenderingNovel View Synthesis

Symbolic Relational Deep Reinforcement Learning based on Graph Neural Networks and Autoregressive Policy Decomposition

2020-09-25 · Jaromír Janisch, Tomáš Pevný, Viliam Lisý

We focus on reinforcement learning (RL) in relational problems that are naturally defined in terms of objects, their relations, and object-centric actions. These problems are characterized by variable state and action sp…

Deep Reinforcement Learningreinforcement-learningReinforcement Learning (RL)Zero-shot Generalization

PRISM: Predictive Recomposition via Semantic Latent Decomposition for View-invariant Video Representation Learning

2026-08-31 · Youngchae Chee, Hosu Lee, Sungjune Park, Junho Kim 외 arxiv

Cross-view video representation learning aims to capture viewpoint-invariant action semantics despite substantial appearance changes across egocentric and exocentric videos. However, existing methods encode each video as…

Representation Learning

Hyperspectral Image Super-resolution via Deep Progressive Zero-centric Residual Learning

2020-06-18 · Zhiyu Zhu, Junhui Hou, Jie Chen, Huanqiang Zeng 외

This paper explores the problem of hyperspectral image (HSI) super-resolution that merges a low resolution HSI (LR-HSI) and a high resolution multispectral image (HR-MSI). The cross-modality distribution of the spatial a…

Hyperspectral Image Super-ResolutionHyperspectral UnmixingImage Super-ResolutionSuper-Resolution