paper-with-me

Papers

Uncertainty-Weighted Image-Event Multimodal Fusion for Video Anomaly Detection

2025-05-05 · Sungheon Jeong, Jihong Park, Mohsen Imani

Most existing video anomaly detectors rely solely on RGB frames, which lack the temporal resolution needed to capture abrupt or transient motion cues, key indicators of anomalous events. To address this limitation, we propose Image-Event Fusion for Video Anomaly Detection (IEF-VAD), a framework that synthesizes event representations directly from RGB videos and fuses them with image features through a principled, uncertainty-aware process. The system (i) models heavy-tailed sensor noise with a Student`s-t likelihood, deriving value-level inverse-variance weights via a Laplace approximation; (ii) applies Kalman-style frame-wise updates to balance modalities over time; and (iii) iteratively refines the fused latent state to erase residual cross-modal noise. Without any dedicated event sensor or frame-level labels, IEF-VAD sets a new state of the art across multiple real-world anomaly detection benchmarks. These findings highlight the utility of synthetic event representations in emphasizing motion cues that are often underrepresented in RGB frames, enabling accurate and robust video understanding across diverse applications without requiring dedicated event sensors. Code and models are available at https://github.com/EavnJeong/IEF-VAD.

📄 PDF Abstract BibTeX arXiv:2505.02393

Code (1)

eavnjeong/ief-vad 공식 구현 pytorch

Tasks

Anomaly DetectionAnomaly Detection In Surveillance VideosVideo Anomaly DetectionVideo Understanding

Similar Papers 제목 키워드 기반

An Attention-based Multi-Scale Feature Learning Network for Multimodal Medical Image Fusion

2022-12-09 · Meng Zhou, Xiaolan Xu, Yuxuan Zhang

Medical images play an important role in clinical applications. Multimodal medical images could provide rich information about patients for physicians to diagnose. The image fusion technique is able to synthesize complem…

Diagnostic

Hyperdimensional Uncertainty Quantification for Multimodal Uncertainty Fusion in Autonomous Vehicles Perception

2025-03-25 · CVPR 2025 1 · Luke Chen, Junyao Wang, Trier Mortlock, Pramod Khargonekar 외

Uncertainty Quantification (UQ) is crucial for ensuring the reliability of machine learning models deployed in real-world autonomous systems. However, existing approaches typically quantify task-level output prediction u…

3D Object DetectionAutonomous Vehiclesobject-detectionObject Detection+2

SEF-MAP: Subspace-Decomposed Expert Fusion for Robust Multimodal HD Map Prediction

2026-02-25 · Haoxiang Fu, Lingfeng Zhang, Hao Li, Ruibing Hu 외 arxiv

High-definition (HD) maps are essential for autonomous driving, yet multi-modal fusion often suffers from inconsistency between camera and LiDAR modalities, leading to performance degradation under low-light conditions, …

Autonomous DrivingPoint Clouds

Evidential Fusion Network for Multimodal Survival Prediction under Missing Modalities

2026-06-18 · Yucheng Xing, Hailan Mo, Zi Wang, Ling Huang 외 arxiv

Recent multimodal survival prediction models have demonstrated strong predictive performance by leveraging complementary information across modalities. However, such models generally assume data completeness and exhibit …

Bridging Modalities: Joint Synthesis and Registration Framework for Aligning Diffusion MRI with T1-Weighted Images

2026-01-16 · Xiaofan Wang, Junyi Wang, Yuqian Chen, Lauren J. O' Donnell 외 arxiv

Multimodal image registration between diffusion MRI (dMRI) and T1-weighted (T1w) MRI images is a critical step for aligning diffusion-weighted imaging (DWI) data with structural anatomical space. Traditional registration…

Image Registration