paper-with-me

Papers

Bidirectional Feature-aligned Motion Transformation for Efficient Dynamic Point Cloud Compression

2025-09-18 · Xuan Deng, Xingtao Wang, Xiandong Meng, Longguang Wang, Tiange Zhang, Xiaopeng Fan, Debin Zhao arxiv

Efficient dynamic point cloud compression (DPCC) critically depends on accurate motion estimation and compensation. However, the inherently irregular structure and substantial local variations of point clouds make this task highly challenging. Existing approaches typically rely on explicit motion estimation, whose encoded motion vectors often fail to capture complex dynamics and inadequately exploit temporal correlations. To address these limitations, we propose a Bidirectional Feature-aligned Motion Transformation (Bi-FMT) framework that implicitly models motion in the feature space. Bi-FMT aligns features across both past and future frames to produce temporally consistent latent representations, which serve as predictive context in a conditional coding pipeline, forming a unified ``Motion + Conditional'' representation. Built upon this bidirectional feature alignment, we introduce a Cross-Transformer Refinement module (CTR) at the decoder side to adaptively refine locally aligned features. By modeling cross-frame dependencies with vector attention, CRT enhances local consistency and restores fine-grained spatial details that are often lost during motion alignment. Moreover, we design a Random Access (RA) reference strategy that treats the bidirectionally aligned features as conditional context, enabling frame-level parallel compression and eliminating the sequential encoding. Extensive experiments demonstrate that Bi-FMT surpasses D-DPCC and AdaDPCC in both compression efficiency and runtime, achieving BD-Rate reductions of 20% (D1) and 9.4% (D1), respectively.

📄 PDF Abstract BibTeX arXiv:2509.14591

Code (0)

등록된 구현이 없습니다.

Tasks

Point Clouds

Similar Papers 제목 키워드 기반

Bidirectional Cross-Modal Prompting for Event-Frame Asymmetric Stereo

2026-04-16 · Ninghui Xu, Fabio Tosi, Lihui Wang, Jiawei Han 외 arxiv

Conventional frame-based cameras capture rich contextual information but suffer from limited temporal resolution and motion blur in dynamic scenes. Event cameras offer an alternative visual representation with higher dyn…

BSAFusion: A Bidirectional Stepwise Feature Alignment Network for Unaligned Medical Image Fusion

2024-12-11 · Huafeng Li, Dayong Su, Qing Cai, Yafei Zhang

If unaligned multimodal medical images can be simultaneously aligned and fused using a single-stage approach within a unified processing framework, it will not only achieve mutual promotion of dual tasks but also help re…

Bidirectional Autoregessive Diffusion Model for Dance Generation

2024-01-01 · CVPR 2024 1 · Canyu Zhang, YouBao Tang, Ning Zhang, Ruei-Sung Lin 외

Dance serves as a powerful medium for expressing human emotions but the lifelike generation of dance is still a considerable challenge. Recently diffusion models have showcased remarkable generative abilities across …

modelMotion Generation

Bidirectional Autoregressive Diffusion Model for Dance Generation

2024-02-06 · Canyu Zhang, YouBao Tang, Ning Zhang, Ruei-Sung Lin 외

Dance serves as a powerful medium for expressing human emotions, but the lifelike generation of dance is still a considerable challenge. Recently, diffusion models have showcased remarkable generative abilities across va…

modelMotion Generation

MT-VAE: Learning Motion Transformations to Generate Multimodal Human Dynamics

2018-08-14 · ECCV 2018 9 · Xinchen Yan, Akash Rastogi, Ruben Villegas, Kalyan Sunkavalli 외

Long-term human motion can be represented as a series of motion modes---motion sequences that capture short-term temporal dynamics---with transitions between them. We leverage this structure and present a novel Motion Tr…

Human DynamicsHuman Pose Forecastingmotion prediction