Superevents: Towards Native Semantic Segmentation for Event-based Cameras
Most successful computer vision models transform low-level features, such as Gabor filter responses, into richer representations of intermediate or mid-level complexity for downstream visual tasks. These mid-level representations have not been explored for event cameras, although it is especially relevant to the visually sparse and often disjoint spatial information in the event stream. By making use of locally consistent intermediate representations, termed as superevents, numerous visual tasks ranging from semantic segmentation, visual tracking, depth estimation shall benefit. In essence, superevents are perceptually consistent local units that delineate parts of an object in a scene. Inspired by recent deep learning architectures, we present a novel method that employs lifetime augmentation for obtaining an event stream representation that is fed to a fully convolutional network to extract superevents. Our qualitative and quantitative experimental results on several sequences of a benchmark dataset highlights the significant potential for event-based downstream applications.
Code (0)
등록된 구현이 없습니다.
Tasks
Depth EstimationSemantic SegmentationVisual TrackingSimilar Papers 제목 키워드 기반
EventSSEG: Event-driven Self-Supervised Segmentation with Probabilistic Attention
Road segmentation is pivotal for autonomous vehicles, yet achieving low latency and low compute solutions using frame based cameras remains a challenge. Event cameras offer a promising alternative. To leverage their low …
Autonomous VehiclesRoad SegmentationOVOSE: Open-Vocabulary Semantic Segmentation in Event-Based Cameras
Event cameras, known for low-latency operation and superior performance in challenging lighting conditions, are suitable for sensitive computer vision tasks such as semantic segmentation in autonomous driving. However, c…
Autonomous DrivingDomain AdaptationKnowledge DistillationOpen Vocabulary Semantic Segmentation+4EV-SegNet: Semantic Segmentation for Event-based Cameras
Event cameras, or Dynamic Vision Sensor (DVS), are very promising sensors which have shown several advantages over frame based cameras. However, most recent work on real applications of these cameras is focused on 3D rec…
3D ReconstructionSegmentationSemantic SegmentationvalidESS: Learning Event-based Semantic Segmentation from Still Images
Retrieving accurate semantic information in challenging high dynamic range (HDR) and high-speed conditions remains an open challenge for image-based algorithms due to severe image degradations. Event cameras promise to a…
Domain AdaptationEvent-based Object SegmentationSegmentationSemantic Segmentation+1CMDA: Cross-Modality Domain Adaptation for Nighttime Semantic Segmentation
Most nighttime semantic segmentation studies are based on domain adaptation approaches and image input. However, limited by the low dynamic range of conventional cameras, images fail to capture structural details and bou…
Domain AdaptationSegmentationSemantic Segmentation