Incremental Human-Object Interaction Detection with Invariant Relation Representation Learning
In open-world environments, human-object interactions (HOIs) evolve continuously, challenging conventional closed-world HOI detection models. Inspired by humans' ability to progressively acquire knowledge, we explore incremental HOI detection (IHOID) to develop agents capable of discerning human-object relations in such dynamic environments. This setup confronts not only the common issue of catastrophic forgetting in incremental learning but also distinct challenges posed by interaction drift and detecting zero-shot HOI combinations with sequentially arriving data. Therefore, we propose a novel exemplar-free incremental relation distillation (IRD) framework. IRD decouples the learning of objects and relations, and introduces two unique distillation losses for learning invariant relation features across different HOI combinations that share the same relation. Extensive experiments on HICO-DET and V-COCO datasets demonstrate the superiority of our method over state-of-the-art baselines in mitigating forgetting, strengthening robustness against interaction drift, and generalization on zero-shot HOIs. Code is available at \href{https://github.com/weiyana/ContinualHOI}{this HTTP URL}
Code (0)
등록된 구현이 없습니다.
Tasks
Human-Object Interaction DetectionRepresentation LearningIncremental LearningSimilar Papers 제목 키워드 기반
Progressively Parsing Interactional Objects for Fine Grained Action Detection
Fine grained video action analysis often requires reliable detection and tracking of various interacting objects and human body parts, denoted as interactional object parsing. However, most of the previous methods based …
Action AnalysisAction DetectionAction RecognitionFine-Grained Action Detection+4Incremental Learning for Robot Perception through HRI
Scene understanding and object recognition is a difficult to achieve yet crucial skill for robots. Recently, Convolutional Neural Networks (CNN), have shown success in this task. However, there is still a gap between the…
Incremental LearningObjectobject-detectionObject Detection+2Online Illumination Invariant Moving Object Detection by Generative Neural Network
Moving object detection (MOD) is a significant problem in computer vision that has many real world applications. Different categories of methods have been proposed to solve MOD. One of the challenges is to separate movin…
Moving Object Detectionobject-detectionObject DetectionLearning 2D Invariant Affordance Knowledge for 3D Affordance Grounding
3D Object Affordance Grounding aims to predict the functional regions on a 3D object and has laid the foundation for a wide range of applications in robotics. Recent advances tackle this problem via learning a mapping be…
Human-Object Interaction DetectionObjectNovel Human-Object Interaction Detection via Adversarial Domain Generalization
We study in this paper the problem of novel human-object interaction (HOI) detection, aiming at improving the generalization ability of the model to unseen scenarios. The challenge mainly stems from the large composition…
Domain GeneralizationHuman-Object Interaction DetectionObjectTriplet