paper-with-me

Papers

DecomPose: Disentangling Cross-Category Optimization Contention for Category-Level 6D Object Pose Estimation

2026-05-15 · Yifan Gao, Lu Zou, Zhangjin Huang, Guoping Wang arxiv

Category-level 6D object pose estimation is typically formulated as a multi-category joint learning problem with fully shared model parameters. However, pronounced geometric heterogeneity across categories entangles incompatible optimization signals in shared modules, resulting in gradient conflicts and negative transfer during training. To address this challenge, we first introduce gradient-based diagnostics to quantify module-level cross-category contention. Building on results of diagnostics, we propose DecomPose, a difficulty-aware decomposition framework that mitigates optimization contention via: (1) difficulty-aware gradient decoupling, which groups categories using a data-driven difficulty proxy and routes each instance to a group-specific correspondence branch to isolate incompatible updates; and (2) stability-driven asymmetric branching, which assigns higher-capacity branches to structurally simple categories as stable optimization anchors while constraining complex categories with lightweight branches to suppress noisy updates and alleviate negative transfer. Extensive experiments on REAL275, CAMERA25, and HouseCat6D demonstrate that DecomPose effectively reduces cross-category optimization contention and delivers superior pose estimation performance across multiple benchmarks.

📄 PDF Abstract BibTeX arXiv:2605.15728

Code (0)

등록된 구현이 없습니다.

Tasks

Pose Estimation

Similar Papers 제목 키워드 기반

Reconstructing Animatable Categories from Videos

2023-05-10 · CVPR 2023 1 · Gengshan Yang, Chaoyang Wang, N Dinesh Reddy, Deva Ramanan

Building animatable 3D models is challenging due to the need for 3D scans, laborious registration, and manual rigging, which are difficult to scale to arbitrary categories. Recently, differentiable rendering provides a p…

3D Shape Reconstruction from VideosDynamic ReconstructionMonocular Reconstruction

CIGMO: Learning categorical invariant deep generative models from grouped data

2021-01-01 · Haruo Hosoya

Images of general objects are often composed of three hidden factors: category (e.g., car or chair), shape (e.g., particular car form), and view (e.g., 3D orientation). While many existing disentangling models can disc…

ClusteringDiversity

Towards a Definition of Disentangled Representations

2018-12-05 · Irina Higgins, David Amos, David Pfau, Sebastien Racaniere 외

How can intelligent agents solve a diverse set of tasks in a data-efficient manner? The disentangled representation learning approach posits that such an agent would benefit from separating out (disentangling) the underl…

Representation Learning

Contention Window Optimization in IEEE 802.11ax Networks with Deep Reinforcement Learning

2020-03-03 · Witold Wydmański, Szymon Szott

The proper setting of contention window (CW) values has a significant impact on the efficiency of Wi-Fi networks. Unfortunately, the standard method used by 802.11 networks is not scalable enough to maintain stable throu…

Deep Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Graph-Structured Deep Learning Framework for Multi-task Contention Identification with High-dimensional Metrics

2026-01-28 · Xiao Yang, Yinan Ni, Yuqi Tang, Zhimin Qiu 외 arxiv

This study addresses the challenge of accurately identifying multi-task contention types in high-dimensional system environments and proposes a unified contention classification framework that integrates representation t…