paper-with-me

홈 › Papers

PEOD: A Pixel-Aligned Event-RGB Benchmark for Object Detection under Challenging Conditions

2025-11-11 · Luoping Cui, Hanqing Liu, Mingjie Liu, Endian Lin, Donghong Jiang, Yuhao Wang, Chuang Zhu arxiv

Robust object detection for challenging scenarios increasingly relies on event cameras, yet existing Event-RGB datasets remain constrained by sparse coverage of extreme conditions and low spatial resolution (<= 640 x 480), which prevents comprehensive evaluation of detectors under challenging scenarios. To address these limitations, we propose PEOD, the first large-scale, pixel-aligned and high-resolution (1280 x 720) Event-RGB dataset for object detection under challenge conditions. PEOD contains 130+ spatiotemporal-aligned sequences and 340k manual bounding boxes, with 57% of data captured under low-light, overexposure, and high-speed motion. Furthermore, we benchmark 14 methods across three input configurations (Event-based, RGB-based, and Event-RGB fusion) on PEOD. On the full test set and normal subset, fusion-based models achieve the excellent performance. However, in illumination challenge subset, the top event-based model outperforms all fusion models, while fusion models still outperform their RGB-based counterparts, indicating limits of existing fusion methods when the frame modality is severely degraded. PEOD establishes a realistic, high-quality benchmark for multimodal perception and facilitates future research.

📄 PDF Abstract BibTeX arXiv:2511.08140

Code (0)

등록된 구현이 없습니다.

Tasks

Robust Object Detection

Similar Papers 제목 키워드 기반

RE-VLM: Event-Augmented Vision-Language Model for Scene Understanding

2026-05-19 · Hanqing Liu, Mingjie Liu, Luoping Cui, Endian Lin 외 arxiv

Conventional vision-language models (VLMs) struggle to interpret scenes captured under adverse conditions (e.g., low light, high dynamic range, or fast motion) because standard RGB images degrade in such environments. Ev…

Scene Understanding

Non-Coaxial Event-Guided Motion Deblurring with Spatial Alignment

2023-01-01 · ICCV 2023 1 · Hoonhee Cho, Yuhwan Jeong, Taewoo Kim, Kuk-Jin Yoon

Motion deblurring from a blurred image is a challenging computer vision problem because frame-based cameras lose information during the blurring process. Several attempts have compensated for the loss of motion infor…

DeblurringImage Deblurring

SIS-Challenge: Event-based Spatio-temporal Instance Segmentation Challenge at the CVPR 2025 Event-based Vision Workshop

2025-08-18 · Friedhelm Hamann, Emil Mededovic, Fabian Gülhan, Yuli Wu 외 arxiv

We present an overview of the Spatio-temporal Instance Segmentation (SIS) challenge held in conjunction with the CVPR 2025 Event-based Vision Workshop. The task is to predict accurate pixel-level segmentation masks of de…

Instance SegmentationEvent-based vision

Object-centric Cross-modal Feature Distillation for Event-based Object Detection

2023-11-09 · Lei LI, Alexander Liniger, Mario Millhaeusler, Vagia Tsiminaki 외

Event cameras are gaining popularity due to their unique properties, such as their low latency and high dynamic range. One task where these benefits can be crucial is real-time object detection. However, RGB detectors st…

Knowledge DistillationObjectobject-detectionObject Detection+1

World Tracing: Generative Pixel-Aligned Geometry Beyond the Visible

2026-06-11 · Hao Zhang, Mohamed El Banani, Jen-Hao Cheng, Paul Zhang 외 arxiv

Image-to-3D methods often trade off faithfulness and completeness: depth estimators are anchored to input pixels but stop at the visible surface, while image-to-3D models generate complete shapes that are often misaligne…

3D scene Editing