paper-with-me

Papers

EvPlug: Learn a Plug-and-Play Module for Event and Image Fusion

2023-12-28 · Jianping Jiang, Xinyu Zhou, Peiqi Duan, Boxin Shi

Event cameras and RGB cameras exhibit complementary characteristics in imaging: the former possesses high dynamic range (HDR) and high temporal resolution, while the latter provides rich texture and color information. This makes the integration of event cameras into middle- and high-level RGB-based vision tasks highly promising. However, challenges arise in multi-modal fusion, data annotation, and model architecture design. In this paper, we propose EvPlug, which learns a plug-and-play event and image fusion module from the supervision of the existing RGB-based model. The learned fusion module integrates event streams with image features in the form of a plug-in, endowing the RGB-based model to be robust to HDR and fast motion scenes while enabling high temporal resolution inference. Our method only requires unlabeled event-image pairs (no pixel-wise alignment required) and does not alter the structure or weights of the RGB-based model. We demonstrate the superiority of EvPlug in several vision tasks such as object detection, semantic segmentation, and 3D hand pose estimation

📄 PDF Abstract BibTeX arXiv:2312.16933

Code (0)

등록된 구현이 없습니다.

Tasks

3D Hand Pose EstimationHand Pose Estimationobject-detectionObject DetectionPose EstimationSemantic Segmentation

Similar Papers 제목 키워드 기반

A Plug-and-Play Bregman ADMM Module for Inferring Event Branches in Temporal Point Processes

2025-01-08 · Qingmei Wang, Yuxin Wu, Yujie Long, Jing Huang 외

An event sequence generated by a temporal point process is often associated with a hidden and structured event branching process that captures the triggering relations between its historical and current events. In this s…

Point Processes

Multi-Focus Temporal Shifting for Precise Event Spotting in Sports Videos

2025-07-10 · Hao Xu, Xinyu Wei, Sam Wells, Sunil Aryal arxiv

Precise Event Spotting (PES) in sports videos requires frame-level recognition of fine-grained actions from single-camera footage. Existing PES models typically incorporate lightweight temporal modules such as the Gate S…

PSTTS: A Plug-and-Play Token Selector for Efficient Event-based Spatio-temporal Representation Learning

2025-09-26 · Xiangmo Zhao, Nan Yang, Yang Wang, Zhanwen Liu arxiv

Mainstream event-based spatio-temporal representation learning methods typically process event streams by converting them into sequences of event frames, achieving remarkable performance. However, they neglect the high s…

Representation Learning

KB-Plugin: A Plug-and-play Framework for Large Language Models to Induce Programs over Low-resourced Knowledge Bases

2024-02-02 · Jiajie Zhang, Shulin Cao, Linmei Hu, Ling Feng 외

Program induction (PI) has become a promising paradigm for using knowledge bases (KBs) to help large language models (LLMs) answer complex knowledge-intensive questions. Nonetheless, PI typically relies on a large number…

Program inductionSelf-Supervised Learning

Attend to the Right Context: A Plug-and-Play Module for Content-Controllable Summarization

2022-12-21 · Wen Xiao, Lesly Miculicich, Yang Liu, Pengcheng He 외

Content-Controllable Summarization generates summaries focused on the given controlling signals. Due to the lack of large-scale training corpora for the task, we propose a plug-and-play module RelAttn to adapt any genera…