paper-with-me

홈 › Papers

Text2LiDAR: Text-guided LiDAR Point Cloud Generation via Equirectangular Transformer

2024-07-29 · Yang Wu, Kaihua Zhang, Jianjun Qian, Jin Xie, Jian Yang

The complex traffic environment and various weather conditions make the collection of LiDAR data expensive and challenging. Achieving high-quality and controllable LiDAR data generation is urgently needed, controlling with text is a common practice, but there is little research in this field. To this end, we propose Text2LiDAR, the first efficient, diverse, and text-controllable LiDAR data generation model. Specifically, we design an equirectangular transformer architecture, utilizing the designed equirectangular attention to capture LiDAR features in a manner with data characteristics. Then, we design a control-signal embedding injector to efficiently integrate control signals through the global-to-focused attention mechanism. Additionally, we devise a frequency modulator to assist the model in recovering high-frequency details, ensuring the clarity of the generated point cloud. To foster development in the field and optimize text-controlled generation performance, we construct nuLiDARtext which offers diverse text descriptors for 34,149 LiDAR point clouds from 850 scenes. Experiments on uncontrolled and text-controlled generation in various forms on KITTI-360 and nuScenes datasets demonstrate the superiority of our approach.

📄 PDF Abstract BibTeX arXiv:2407.19628

Code (0)

등록된 구현이 없습니다.

Tasks

Point Cloud Generation

Methods 이 논문이 사용한 방법론

Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Attention 설명 없음

Similar Papers 제목 키워드 기반

MGTANet: Encoding Sequential LiDAR Points Using Long Short-Term Motion-Guided Temporal Attention for 3D Object Detection

2022-12-01 · Junho Koh, Junhyung Lee, Youngwoo Lee, Jaekyum Kim 외

Most scanning LiDAR sensors generate a sequence of point clouds in real-time. While conventional 3D object detectors use a set of unordered LiDAR points acquired over a fixed time interval, recent studies have revealed t…

3D Object DetectionObjectobject-detectionObject Detection

Leveraging Sparse LiDAR for RAFT-Stereo: A Depth Pre-Fill Perspective

2025-07-26 · Jinsu Yoo, Sooyoung Jeon, Zanming Huang, Tai-Yu Pan 외 arxiv

We investigate LiDAR guidance within the RAFT-Stereo framework, aiming to improve stereo matching accuracy by injecting precise LiDAR depth into the initial disparity map. We find that the effectiveness of LiDAR guidance…

LiDAR-PTQ: Post-Training Quantization for Point Cloud 3D Object Detection

2024-01-29 · Sifan Zhou, Liang Li, Xinyu Zhang, Bo Zhang 외

Due to highly constrained computing power and memory, deploying 3D lidar-based detectors on edge devices equipped in autonomous vehicles and robots poses a crucial challenge. Being a convenient and straightforward model …

3D Object DetectionAutonomous VehiclesModel Compressionobject-detection+2

Attention-Guided Lidar Segmentation and Odometry Using Image-to-Point Cloud Saliency Transfer

2023-08-28 · Guanqun Ding, Nevrez Imamoglu, Ali Caglayan, Masahiro Murakawa 외

LiDAR odometry estimation and 3D semantic segmentation are crucial for autonomous driving, which has achieved remarkable advances recently. However, these tasks are challenging due to the imbalance of points in different…

3D Semantic SegmentationAutonomous DrivingSegmentationSemantic Segmentation+1

LidarCLIP or: How I Learned to Talk to Point Clouds

2022-12-13 · Georg Hess, Adam Tonderski, Christoffer Petersson, Kalle Åström 외

Research connecting text and images has recently seen several breakthroughs, with models like CLIP, DALL-E 2, and Stable Diffusion. However, the connection between text and other visual modalities, such as lidar data, ha…

Image GenerationRetrievalzero-shot-classificationZero-Shot Learning