paper-with-me

홈 › Papers

PoseStreamer: A Multi-modal Framework for 3D Tracking of Unseen Moving Objects

2025-12-28 · Huiming Yang, Linglin Liao, Fei Ding, Sibo Wang, Zijian Zeng arxiv

Six degree of freedom (6DoF) pose estimation for novel objects is a critical task in computer vision, yet it faces significant challenges in high-speed and low-light scenarios where standard RGB cameras suffer from motion blur. While event cameras offer a promising solution due to their high temporal resolution, current 6DoF pose estimation methods typically yield suboptimal performance in high-speed object moving scenarios. To address this gap, we propose PoseStreamer, a robust multi-modal 6DoF pose estimation framework designed specifically on high-speed moving scenarios. Our approach integrates three core components: an Adaptive Pose Memory Queue that utilizes historical orientation cues for temporal consistency, an Object-centric 2D Tracker that provides strong 2D priors to boost 3D center recall, and a Ray Pose Filter for geometric refinement along camera rays. Furthermore, we introduce MoCapCube6D, a novel multi-modal dataset constructed to benchmark performance under rapid motion. Extensive experiments demonstrate that PoseStreamer not only achieves superior accuracy in high-speed moving scenarios, but also exhibits strong generalizability as a template-free framework for unseen moving objects.

📄 PDF Abstract BibTeX arXiv:2512.22979

Code (0)

등록된 구현이 없습니다.

Tasks

Pose Estimation

Similar Papers 제목 키워드 기반

OVTR: End-to-End Open-Vocabulary Multiple Object Tracking with Transformer

2025-03-13 · Jinyang Li, En Yu, Sijia Chen, Wenbing Tao

Open-vocabulary multiple object tracking aims to generalize trackers to unseen categories during training, enabling their application across a variety of real-world scenarios. However, the existing open-vocabulary tracke…

Decodermultimodal interactionMultiple Object TrackingMultiple Object Tracking with Transformer+1

A Tri-Modal Dataset and a Baseline System for Tracking Unmanned Aerial Vehicles

2025-11-23 · Tianyang Xu, Jinjie Gu, Xuefeng Zhu, XiaoJun Wu 외 arxiv

With the proliferation of low altitude unmanned aerial vehicles (UAVs), visual multi-object tracking is becoming a critical security technology, demanding significant robustness even in complex environmental conditions. …

Multi-Object Tracking

Unified Sequence-to-Sequence Learning for Single- and Multi-Modal Visual Object Tracking

2023-04-27 · CVPR 2023 1 · Xin Chen, Ben Kang, Jiawen Zhu, Dong Wang 외

In this paper, we introduce a new sequence-to-sequence learning framework for RGB-based and multi-modal object tracking. First, we present SeqTrack for RGB-based tracking. It casts visual tracking as a sequence generatio…

DecoderObjectObject TrackingRgb-T Tracking+2

M3imic: Learning a Versatile Whole-Body Controller for Multimodal Motion Mimicking

2026-06-03 · Zuxing Lu, Ziang Zheng, Yao Lyu, Jingyu Liu 외 arxiv

Building a general-purpose whole-body controller is essential for enabling diverse motion capabilities in humanoid robots across a wide range of downstream tasks, including locomotion and loco-manipulation. Different tas…

Reinforcement Learning

Open3DTrack: Towards Open-Vocabulary 3D Multi-Object Tracking

2024-10-02 · Ayesha Ishaq, Mohamed El Amine Boudjoghra, Jean Lahoud, Fahad Shahbaz Khan 외

3D multi-object tracking plays a critical role in autonomous driving by enabling the real-time monitoring and prediction of multiple objects' movements. Traditional 3D tracking systems are typically constrained by predef…

3D Multi-Object TrackingAutonomous DrivingMulti-Object TrackingObject+1