paper-with-me

Papers

On the Geometry of Learned Representations in Event-Based Multi-Modal Egomotion Estimation

2026-07-17 · Stefano Silvestrini, Michele Ceresoli arxiv

Classical approaches to event-based egomotion estimation, including those adopted by the top-performing teams of the ELOPE challenge, rely on geometric optimization frameworks such as contrast maximization, homography estimation, or dense optical flow combined with analytic motion inversion. This work investigates the geometric structure that emerges inside a multi-modal network for egomotion estimation. Event tensors, inertial measurements, and range signals are fused through a cross-modal attention architecture and trained in a batch setting. We analyze the latent space geometry and attention dynamics, showing that (i) embeddings lie on low-dimensional manifolds aligned with motion variables, (ii) attention weights adapt with angular excitation and visual reliability, and (iii) the fused representation recovers classical observability cues. These results bridge analytical estimation theory and modern data-driven fusion.

📄 PDF Abstract BibTeX arXiv:2607.15794

Code (0)

등록된 구현이 없습니다.

Tasks

Homography Estimation

Similar Papers 제목 키워드 기반

Multimodal Sparse Coding for Event Detection

2016-05-17 · Youngjune Gwon, William Campbell, Kevin Brady, Douglas Sturim 외

Unsupervised feature learning methods have proven effective for classification tasks based on a single modality. We present multimodal sparse coding for learning feature representations shared across multiple modalities.…

ClassificationEvent DetectionGeneral Classification

Diverse via bounded Agreement: Geometric Regularization for Multimodal Fusion

2026-01-29 · Zixuan Xia, Hao Wang, Pengcheng Weng, Yanyu Qian 외 arxiv

Multimodal fusion is often treated as an optimization-balancing problem, where training signals are adjusted to prevent one modality from dominating the others. However, balanced optimization does not fully determine the…

Representation Learning

Efficient Multi-Timescale Event Representations for Feed-Forward Object Detection

2026-09-04 · Fredrik Lundell, Per-Erik Forssen, Mårten Wadenbäck, Astrid Lundmark arxiv

Autonomous systems require robust low-latency perception under rapidly changing scene dynamics and challenging illumination. In event cameras object detection commonly relies on recurrent architectures to accumulate spar…

Object Detection

xModel-KD: Cross-modal Knowledge Distillation for 3D Scene Perception using LiDAR

2026-05-28 · Thenukan Pathmanathan, Kanchan Keisham, Thangarajah Akilan arxiv

Point cloud segmentation is a fundamental task in 3D scene understanding. Its progress is constrained by the high cost and time required for dense 3D annotations, making labeled samples difficult to obtain. Beyond annota…

Point Cloud SegmentationKnowledge DistillationScene UnderstandingPoint Clouds

EventFace: Event-Based Face Recognition via Structure-Driven Spatiotemporal Modeling

2026-04-08 · Qingguo Meng, Xingbo Dong, Zhe Jin, Massimo Tistarelli arxiv

Event cameras offer a promising sensing modality for face recognition due to their inherent advantages in illumination robustness and privacy-friendliness. However, because event streams lack the stable photometric appea…

Face Recognition