paper-with-me

홈 › Papers

CASA: Class-Agnostic Shared Attributes in Vision-Language Models for Efficient Incremental Object Detection

2024-10-08 · Mingyi Guo, Yuyang Liu, Zhiyuan Yan, Zongying Lin, Peixi Peng, Yonghong Tian

Incremental object detection is fundamentally challenged by catastrophic forgetting. A major factor contributing to this issue is background shift, where background categories in sequential tasks may overlap with either previously learned or future unseen classes. To address this, we propose a novel method called Class-Agnostic Shared Attribute Base (CASA) that encourages the model to learn category-agnostic attributes shared across incremental classes. Our approach leverages an LLM to generate candidate textual attributes, selects the most relevant ones based on the current training data, and records their importance in an assignment matrix. For subsequent tasks, the retained attributes are frozen, and new attributes are selected from the remaining candidates, ensuring both knowledge retention and adaptability. Extensive experiments on the COCO dataset demonstrate the state-of-the-art performance of our method.

📄 PDF Abstract BibTeX arXiv:2410.05804

Code (0)

등록된 구현이 없습니다.

Tasks

Attributeobject-detectionObject Detectionparameter-efficient fine-tuning

Methods 이 논문이 사용한 방법론

CLIP Contrastive Language-Image Pre-training (CLIP), consisting of a simplified version of ConVIRT trained from scratch, is an efficient method of image representation learning…
BASE 설명 없음

Similar Papers 제목 키워드 기반

CASA: Category-agnostic Skeletal Animal Reconstruction

2022-11-04 · Yuefan Wu, Zeyuan Chen, Shaowei Liu, Zhongzheng Ren 외

Recovering the skeletal shape of an animal from a monocular video is a longstanding challenge. Prevailing animal reconstruction methods often adopt a control-point driven animation model and optimize bone transforms indi…

Retrieval

Z-1: Efficient Reinforcement Learning for Vision-Language-Action Models

2026-06-30 · Lang Cao, Renhong Chen, Luyi Li, Peng Wang 외 arxiv

Vision-Language-Action (VLA) models offer a promising framework for robotic manipulation by connecting language instructions, visual observations, and continuous control. However, most existing policies remain limited by…

Reinforcement LearningContinuous Control

Enhancing Terrestrial Net Primary Productivity Estimation with EXP-CASA: A Novel Light Use Efficiency Model Approach

2024-06-28 · Guanzhou Chen, Kaiqi Zhang, Xiaodong Zhang, Hong Xie 외

The Light Use Efficiency model, epitomized by the CASA model, is extensively applied in the quantitative estimation of vegetation Net Primary Productivity. However, the classic CASA model is marked by significant complex…

Hallucination reduction with CASAL: Contrastive Activation Steering For Amortized Learning

2025-09-25 · Wannan, Yang, Xinchi Qiu, Lei Yu 외 arxiv

Large Language Models (LLMs) exhibit impressive capabilities but often hallucinate, confidently providing incorrect answers instead of admitting ignorance. Prior work has shown that models encode linear representations o…

What Are We Actually Benchmarking in Robot Manipulation?

2026-06-02 · Tianchong Jiang, Xiangshan Tan, Samuel Wheeler, Luzhe Sun 외 arxiv

A robotics benchmark score measures success under one fixed evaluation setup, yet is routinely treated as evidence of general manipulation capability. We identify four failure modes, each of which weakens or invalidates …

Robot Manipulation