paper-with-me

Papers

FUSE: Label-Free Image-Event Joint Monocular Depth Estimation via Frequency-Decoupled Alignment and Degradation-Robust Fusion

2025-03-25 · Pihai Sun, Junjun Jiang, Yuanqi Yao, Youyu Chen, Wenbo Zhao, Kui Jiang, Xianming Liu

Image-event joint depth estimation methods leverage complementary modalities for robust perception, yet face challenges in generalizability stemming from two factors: 1) limited annotated image-event-depth datasets causing insufficient cross-modal supervision, and 2) inherent frequency mismatches between static images and dynamic event streams with distinct spatiotemporal patterns, leading to ineffective feature fusion. To address this dual challenge, we propose Frequency-decoupled Unified Self-supervised Encoder (FUSE) with two synergistic components: The Parameter-efficient Self-supervised Transfer (PST) establishes cross-modal knowledge transfer through latent space alignment with image foundation models, effectively mitigating data scarcity by enabling joint encoding without depth ground truth. Complementing this, we propose the Frequency-Decoupled Fusion module (FreDFuse) to explicitly decouple high-frequency edge features from low-frequency structural components, resolving modality-specific frequency mismatches through physics-aware fusion. This combined approach enables FUSE to construct a universal image-event encoder that only requires lightweight decoder adaptation for target datasets. Extensive experiments demonstrate state-of-the-art performance with 14% and 24.9% improvements in Abs.Rel on MVSEC and DENSE datasets. The framework exhibits remarkable zero-shot adaptability to challenging scenarios including extreme lighting and motion blur, significantly advancing real-world deployment capabilities. The source code for our method is publicly available at: https://github.com/sunpihai-up/FUSE

📄 PDF Abstract BibTeX arXiv:2503.19739

Code (1)

sunpihai-up/fuse 공식 구현 pytorch

Tasks

Depth EstimationMonocular Depth EstimationTransfer Learning

Similar Papers 제목 키워드 기반

Label-Free Event-based Object Recognition via Joint Learning with Image Reconstruction from Events

2023-08-18 · ICCV 2023 1 · Hoonhee Cho, Hyeonseong Kim, Yujeong Chae, Kuk-Jin Yoon

Recognizing objects from sparse and noisy events becomes extremely difficult when paired images and category labels do not exist. In this paper, we study label-free event-based object recognition where category labels an…

Image ReconstructionObjectObject Recognition

Towards Label-Free Single-Cell Phenotyping Using Multi-Task Learning

2026-05-14 · Saqib Nazir, Ardhendu Behera arxiv

Label-free single-cell imaging offers a scalable, non-invasive alternative to fluorescence-based cytometry, yet inferring molecular phenotypes directly from bright-field morphology remains challenging. We present a unifi…

Multi-Task Learning

EventDance++: Language-guided Unsupervised Source-free Cross-modal Adaptation for Event-based Object Recognition

2024-09-19 · Xu Zheng, Lin Wang

In this paper, we address the challenging problem of cross-modal (image-to-events) adaptation for event-based recognition without accessing any labeled source image data. This task is arduous due to the substantial modal…

Object Recognition

Lifting Monocular Events to 3D Human Poses

2021-04-21 · Gianluca Scarpellini, Pietro Morerio, Alessio Del Bue

This paper presents a novel 3D human pose estimation approach using a single stream of asynchronous events as input. Most of the state-of-the-art approaches solve this task with RGB cameras, however struggling when subje…

3D Human Pose Estimation3D Pose EstimationEvent-based visionPose Estimation

Fused Gromov-Wasserstein Alignment for Hawkes Processes

2019-10-04 · Dixin Luo, Hongteng Xu, Lawrence Carin

We propose a novel fused Gromov-Wasserstein alignment method to jointly learn the Hawkes processes in different event spaces, and align their event types. Given two Hawkes processes, we use fused Gromov-Wasserstein discr…