Driving Scene Perception Network: Real-time Joint Detection, Depth Estimation and Semantic Segmentation
As the demand for enabling high-level autonomous driving has increased in recent years and visual perception is one of the critical features to enable fully autonomous driving, in this paper, we introduce an efficient approach for simultaneous object detection, depth estimation and pixel-level semantic segmentation using a shared convolutional architecture. The proposed network model, which we named Driving Scene Perception Network (DSPNet), uses multi-level feature maps and multi-task learning to improve the accuracy and efficiency of object detection, depth estimation and image segmentation tasks from a single input image. Hence, the resulting network model uses less than 850 MiB of GPU memory and achieves 14.0 fps on NVIDIA GeForce GTX 1080 with a 1024x512 input image, and both precision and efficiency have been improved over combination of single tasks.
Code (0)
등록된 구현이 없습니다.
Tasks
Autonomous DrivingDepth EstimationGPUImage SegmentationMulti-Task Learningobject-detectionObject DetectionSegmentationSemantic SegmentationSimilar Papers 제목 키워드 기반
StreamYOLO: Real-time Object Detection for Streaming Perception
The perceptive models of autonomous driving require fast inference within a low latency for safety. While existing works ignore the inevitable environmental changes after processing, streaming perception jointly evaluate…
Autonomous DrivingObjectobject-detectionObject Detection+1LiDAR-BEVMTN: Real-Time LiDAR Bird's-Eye View Multi-Task Perception Network for Autonomous Driving
LiDAR is crucial for robust 3D scene perception in autonomous driving. LiDAR perception has the largest body of literature after camera perception. However, multi-task learning across tasks like detection, segmentation, …
3D Object DetectionAutonomous DrivingMotion EstimationMotion Segmentation+6A Reliable Context-Aware and Temporal Planning Framework for Autonomous Driving
Safe operation of autonomous vehicles in dense urban traffic depends on perception and planning that remain reliable when onboard sensing is degraded. In real driving conditions, camera observations are frequently corrup…
Computational EfficiencyAutonomous VehiclesScene UnderstandingTrajectory PlanningnuCarla: A nuScenes-Style Bird's-Eye View Perception Dataset for CARLA Simulation
End-to-end (E2E) autonomous driving heavily relies on closed-loop simulation, where perception, planning, and control are jointly trained and evaluated in interactive environments. Yet, most existing datasets are collect…
Autonomous DrivingRadarMP: Motion Perception for 4D mmWave Radar in Autonomous Driving
Accurate 3D scene motion perception significantly enhances the safety and reliability of an autonomous driving system. Benefiting from its all-weather operational capability and unique perceptual properties, 4D mmWave ra…
Point Cloud GenerationAutonomous VehiclesAutonomous Driving