paper-with-me

Papers

RESOLVE: A Multi-Resolution and Multi-Modal Dataset for Roadside Cooperative Perception

2026-06-30 · Shaozu Ding, Linan Song, Marco De Vincenzi, Dajiang Suo arxiv

LiDAR has increasingly been integrated into traffic cameras to expand coverage and mitigate occlusion in roadside cooperative perception. However, how unimodal and camera-LiDAR fusion architectures behave under variations in LiDAR point sparsity induced by sensor configurations and scene-dependent sensing conditions remains underexplored. We introduce RESOLVE, a large-scale real-world benchmark dataset featuring multi-resolution roadside LiDAR and synchronized camera-LiDAR sensing for systematic evaluation of unimodal and fusion-based architectures in roadside 3D detection and tracking. RESOLVE contains over 100k images and 26k point cloud frames with 220k manually annotated bounding boxes, captured at a real-world urban intersection across diverse lighting and weather conditions and spanning 10 classes of traffic participants. In particular, RESOLVE enables controlled evaluation across three LiDAR resolution levels while keeping all other sensing and environmental factors fixed. This allows fair cross-architecture comparisons under point cloud distribution shifts resulting from resolution variations, sensing distance, and training-inference resolution mismatches. Results from extensive benchmark experiments reveal insights into how multimodal fusion can compensate for LiDAR point sparsity, offering clues for designing cost-efficient roadside multimodal perception. The dataset and benchmark codes are available at https://github.com/ASU-Suo-Lab/RESOLVE.

📄 PDF Abstract BibTeX arXiv:2606.31895

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Three-Dimensional, Multimodal Synchrotron Data for Machine Learning Applications

2024-09-11 · Calum Green, Sharif Ahmed, Shashidhara Marathe, Liam Perera 외

Machine learning techniques are being increasingly applied in medical and physical sciences across a variety of imaging modalities; however, an important issue when developing these tools is the availability of good qual…

3D ReconstructionSuper-Resolution

Cross-modal Diffusion Modelling for Super-resolved Spatial Transcriptomics

2024-04-19 · Xiaofei Wang, Xingxu Huang, Stephen J. Price, Chao Li

The recent advancement of spatial transcriptomics (ST) allows to characterize spatial gene expression within tissue for discovery research. However, current ST platforms suffer from low resolution, hindering in-depth und…

Super-Resolution

Multi-Modal Super Resolution for Dense Microscopic Particle Size Estimation

2020-10-19 · Sarvesh Patil, Chava Y P D Phani Rajanish, Naveen Margankunte

Particle Size Analysis (PSA) is an important process carried out in a number of industries, which can significantly influence the properties of the final product. A ubiquitous instrument for this purpose is the Optical M…

object-detectionObject DetectionSuper-ResolutionTranslation

Multimodal Image Super-resolution via Deep Unfolding with Side Information

2019-10-18 · Iman Marivani, Evaggelia Tsiligianni, Bruno Cornelis, Nikos Deligiannis

Deep learning methods have been successfully applied to various computer vision tasks. However, existing neural network architectures do not per se incorporate domain knowledge about the addressed problem, thus, understa…

Image Super-ResolutionSuper-Resolution

GRAVL-BERT: Graphical Visual-Linguistic Representations for Multimodal Coreference Resolution

2022-10-01 · COLING 2022 10 · Danfeng Guo, Arpit Gupta, Sanchit Agarwal, Jiun-Yu Kao 외

Learning from multimodal data has become a popular research topic in recent years. Multimodal coreference resolution (MCR) is an important task in this area. MCR involves resolving the references across different modalit…

coreference-resolutionCoreference ResolutionVisual Grounding