paper-with-me

Papers

SpotNet: An Image Centric, Lidar Anchored Approach To Long Range Perception

2024-05-24 · Louis Foucard, Samar Khanna, Yi Shi, Chi-Kuei Liu, Quinn Z Shen, Thuyen Ngo, Zi-Xiang Xia

In this paper, we propose SpotNet: a fast, single stage, image-centric but LiDAR anchored approach for long range 3D object detection. We demonstrate that our approach to LiDAR/image sensor fusion, combined with the joint learning of 2D and 3D detection tasks, can lead to accurate 3D object detection with very sparse LiDAR support. Unlike more recent bird's-eye-view (BEV) sensor-fusion methods which scale with range $r$ as $O(r^2)$, SpotNet scales as $O(1)$ with range. We argue that such an architecture is ideally suited to leverage each sensor's strength, i.e. semantic understanding from images and accurate range finding from LiDAR data. Finally we show that anchoring detections on LiDAR points removes the need to regress distances, and so the architecture is able to transfer from 2MP to 8MP resolution images without re-training.

📄 PDF Abstract BibTeX arXiv:2405.15843

Code (0)

등록된 구현이 없습니다.

Tasks

3D Object Detectionobject-detectionObject DetectionSensor Fusion

Similar Papers 제목 키워드 기반

SpotNet - Learned iterations for cell detection in image-based immunoassays

2018-10-15 · Pol del Aguila Pla, Vidit Saxena, Joakim Jaldén

Accurate cell detection and counting in the image-based ELISpot and FluoroSpot immunoassays is a challenging task. Methodology recently proposed by our group matches human accuracy by leveraging knowledge of the underlyi…

Cell Detection

BlindSpotNet: Seeing Where We Cannot See

2022-07-08 · Taichi Fukuda, Kotaro Hasegawa, Shinya Ishizaki, Shohei Nobuhara 외

We introduce 2D blind spot estimation as a critical visual task for road scene understanding. By automatically detecting road regions that are occluded from the vehicle's vantage point, we can proactively alert a manual …

Depth EstimationMonocular Depth Estimationroad scene understandingScene Understanding+1

BEVDilation: LiDAR-Centric Multi-Modal Fusion for 3D Object Detection

2025-12-02 · Guowen Zhang, Chenhang He, Liyi Chen, Lei Zhang arxiv

Integrating LiDAR and camera information in the bird's eye view (BEV) representation has demonstrated its effectiveness in 3D object detection. However, because of the fundamental disparity in geometric accuracy between …

Computational Efficiency3D Object DetectionDepth EstimationPoint Clouds

VERIA: Verification-Centric Multimodal Instance Augmentation for Long-Tailed 3D Object Detection

2026-03-25 · Jumin Lee, Siyeong Lee, Namil Kim, Sung-Eui Yoon arxiv

Long-tail distributions in driving datasets pose a fundamental challenge for 3D perception, as rare classes exhibit substantial intra-class diversity yet available samples cover this variation space only sparsely. Existi…

3D Object Detection

EgoGenesis: Egocentric World-Action Modeling with Online Anchored Projective Memory and Action-3D RoPE

2026-07-30 · Zexuan Yan, Yuzhou Wu, Yue Ma, Zonghang He 외 arxiv

Egocentric video offers rich manipulation experience for embodied AI, yet collecting diverse egocentric data across scenes, objects, motions, and embodiments remains costly. We present \method, an egocentric world-action…

Video Generation