paper-with-me

홈 › Papers

Homography Guided Temporal Fusion for Road Line and Marking Segmentation

2024-04-11 · ICCV 2023 1 · Shan Wang, Chuong Nguyen, Jiawei Liu, Kaihao Zhang, Wenhan Luo, Yanhao Zhang, Sundaram Muthu, Fahira Afzal Maken, Hongdong Li

Reliable segmentation of road lines and markings is critical to autonomous driving. Our work is motivated by the observations that road lines and markings are (1) frequently occluded in the presence of moving vehicles, shadow, and glare and (2) highly structured with low intra-class shape variance and overall high appearance consistency. To solve these issues, we propose a Homography Guided Fusion (HomoFusion) module to exploit temporally-adjacent video frames for complementary cues facilitating the correct classification of the partially occluded road lines or markings. To reduce computational complexity, a novel surface normal estimator is proposed to establish spatial correspondences between the sampled frames, allowing the HomoFusion module to perform a pixel-to-pixel attention mechanism in updating the representation of the occluded road lines or markings. Experiments on ApolloScape, a large-scale lane mark segmentation dataset, and ApolloScape Night with artificial simulated night-time road conditions, demonstrate that our method outperforms other existing SOTA lane mark segmentation models with less than 9\% of their parameters and computational complexity. We show that exploiting available camera intrinsic data and ground plane assumption for cross-frame correspondence can lead to a light-weight network with significantly improved performances in speed and accuracy. We also prove the versatility of our HomoFusion approach by applying it to the problem of water puddle segmentation and achieving SOTA performance.

📄 PDF Abstract BibTeX arXiv:2404.07626

Code (1)

shanwang-shan/homofusion 공식 구현 pytorch

Tasks

Autonomous DrivingSegmentation

Methods 이 논문이 사용한 방법론

SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…

Similar Papers 제목 키워드 기반

CodingHomo: Bootstrapping Deep Homography With Video Coding

2025-04-16 · Yike Liu, Haipeng Li, Shuaicheng Liu, Bing Zeng

Homography estimation is a fundamental task in computer vision with applications in diverse fields. Recent advances in deep learning have improved homography estimation, particularly with unsupervised learning approaches…

Homography Estimation

GyroFlow+: Gyroscope-Guided Unsupervised Deep Homography and Optical Flow Learning

2023-01-23 · Haipeng Li, Kunming Luo, Bing Zeng, Shuaicheng Liu

Existing homography and optical flow methods are erroneous in challenging scenes, such as fog, rain, night, and snow because the basic assumptions such as brightness and gradient constancy are broken. To address this iss…

DecoderHomography EstimationOptical Flow Estimation

HOLO: Homography-Guided Pose Estimator Network for Fine-Grained Visual Localization on SD Maps

2026-01-06 · Xuchang Zhong, Xu Cao, Jinke Feng, Hao Fang arxiv

Visual localization on standard-definition (SD) maps has emerged as a promising low-cost and scalable solution for autonomous driving. However, existing regression-based approaches often overlook inherent geometric prior…

Visual LocalizationAutonomous Driving

Zero-Parameter Geometric Gating for Temporally Stable Low-Altitude UAV Video Semantic Segmentation

2026-06-08 · Jingpu Yang, Fengxian Ji, Zhengzhao Lai, Juanfan Wu 외 arxiv

Video semantic segmentation for low-altitude UAVs requires temporal consistency, yet dense optical flow introduces spatially structured noise in the planar regions that dominate aerial imagery. We propose a zero-paramete…

Video Semantic SegmentationSemantic Similarity

TransFill: Reference-guided Image Inpainting by Merging Multiple Color and Spatial Transformations

2021-03-29 · CVPR 2021 1 · Yuqian Zhou, Connelly Barnes, Eli Shechtman, Sohrab Amirghodsi

Image inpainting is the task of plausibly restoring missing pixels within a hole region that is to be removed from a target image. Most existing technologies exploit patch similarities within the image, or leverage large…

Image Inpainting