paper-with-me

홈 › Papers

Borrowing from anything: A generalizable framework for reference-guided instance editing

2025-12-17 · Shengxiao Zhou, Chenghua Li, Jianhao Huang, Qinghao Hu, Yifan Zhang arxiv

Reference-guided instance editing is fundamentally limited by semantic entanglement, where a reference's intrinsic appearance is intertwined with its extrinsic attributes. The key challenge lies in disentangling what information should be borrowed from the reference, and determining how to apply it appropriately to the target. To tackle this challenge, we propose GENIE, a Generalizable Instance Editing framework capable of achieving explicit disentanglement. GENIE first corrects spatial misalignments with a Spatial Alignment Module (SAM). Then, an Adaptive Residual Scaling Module (ARSM) learns what to borrow by amplifying salient intrinsic cues while suppressing extrinsic attributes, while a Progressive Attention Fusion (PAF) mechanism learns how to render this appearance onto the target, preserving its structure. Extensive experiments on the challenging AnyInsertion dataset demonstrate that GENIE achieves state-of-the-art fidelity and robustness, setting a new standard for disentanglement-based instance editing.

📄 PDF Abstract BibTeX arXiv:2512.15138

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

RewardAnything: Generalizable Principle-Following Reward Models

2025-06-04 · Zhuohao Yu, Jiali Zeng, Weizheng Gu, Yidong Wang 외

Reward Models, essential for guiding Large Language Model optimization, are typically trained on fixed preference datasets, resulting in rigid alignment to single, implicit preference distributions. This prevents adaptat…

Instruction FollowingLarge Language ModelModel Optimization

Towards Generalizable Scene Change Detection

2025-01-01 · CVPR 2025 1 · Jae-Woo Kim, Ue-Hwan Kim

While current state-of-the-art Scene Change Detection (SCD) approaches achieve impressive results in well-trained research data, they become unreliable under unseen environments and different temporal conditions; in-…

Change DetectionScene Change Detection

Encoding Structural Constraints into Segment Anything Models via Probabilistic Graphical Models

2025-09-26 · Yu Li, Da Chang, Xi Xiao arxiv

While the Segment Anything Model (SAM) has achieved remarkable success in image segmentation, its direct application to medical imaging remains hindered by fundamental challenges, including ambiguous boundaries, insuffic…

Medical Image Segmentation

MoCapAnything: Unified 3D Motion Capture for Arbitrary Skeletons from Monocular Videos

2025-12-11 · Kehong Gong, Zhengyu Wen, Weixia He, Mingxi Xu 외 arxiv

Motion capture now underpins content creation far beyond digital humans, yet most existing pipelines remain species- or template-specific. We formalize this gap as Category-Agnostic Motion Capture (CAMoCap): given a mono…

Generalizable Visual Reinforcement Learning with Segment Anything Model

2023-12-28 · Ziyu Wang, Yanjie Ze, Yifei Sun, Zhecheng Yuan 외

Learning policies that can generalize to unseen environments is a fundamental challenge in visual reinforcement learning (RL). While most current methods focus on acquiring robust visual representations through auxiliary…

Data Augmentationmodelreinforcement-learningReinforcement Learning+1