paper-with-me

홈 › Papers

LabelAny3D: Label Any Object 3D in the Wild

2026-01-04 · Jin Yao, Radowan Mahmud Redoy, Sebastian Elbaum, Matthew B. Dwyer, Zezhou Cheng arxiv

Detecting objects in 3D space from monocular input is crucial for applications ranging from robotics to scene understanding. Despite advanced performance in the indoor and autonomous driving domains, existing monocular 3D detection models struggle with in-the-wild images due to the lack of 3D in-the-wild datasets and the challenges of 3D annotation. We introduce LabelAny3D, an \emph{analysis-by-synthesis} framework that reconstructs holistic 3D scenes from 2D images to efficiently produce high-quality 3D bounding box annotations. Built on this pipeline, we present COCO3D, a new benchmark for open-vocabulary monocular 3D detection, derived from the MS-COCO dataset and covering a wide range of object categories absent from existing 3D datasets. Experiments show that annotations generated by LabelAny3D improve monocular 3D detection performance across multiple benchmarks, outperforming prior auto-labeling approaches in quality. These results demonstrate the promise of foundation-model-driven annotation for scaling up 3D recognition in realistic, open-world settings.

📄 PDF Abstract BibTeX arXiv:2601.01676

Code (0)

등록된 구현이 없습니다.

Tasks

Scene UnderstandingAutonomous Driving

Similar Papers 제목 키워드 기반

Label Anything: Multi-Class Few-Shot Semantic Segmentation with Visual Prompts

2024-07-02 · Pasquale De Marinis, Nicola Fanelli, Raffaele Scaringi, Emanuele Colonna 외

We present Label Anything, an innovative neural network architecture designed for few-shot semantic segmentation (FSS) that demonstrates remarkable generalizability across multiple classes with minimal examples required …

Few-Shot Semantic SegmentationSemantic Segmentation

Semi-Supervised Domain Adaptation for Wildfire Detection

2024-04-02 · Jooyoung Jang, Youngseo Cha, Jisu Kim, SooHyung Lee 외

Recently, both the frequency and intensity of wildfires have increased worldwide, primarily due to climate change. In this paper, we propose a novel protocol for wildfire detection, leveraging semi-supervised Domain Adap…

Domain Adaptationobject-detectionObject DetectionSemi-supervised Domain Adaptation

Extended Labeled Faces in-the-Wild (ELFW): Augmenting Classes for Face Segmentation

2020-06-24 · Rafael Redondo, Jaume Gibert

Existing face datasets often lack sufficient representation of occluding objects, which can hinder recognition, but also supply meaningful information to understand the visual context. In this work, we introduce Extended…

BenchmarkingData Augmentation

WILD SAM: A Simulated-and-Real Data Augmentation for Autonomous Driving Perception under Challenging Weather

2026-05-01 · Hamed Khatounabadi, Xiaohu Lu, Hayder Radha arxiv

The performance of state-of-the-art object detectors degrades significantly under adverse weather, causing a safety-critical domain shift problem for autonomous vehicles. Recent efforts address this problem by relying on…

Autonomous VehiclesAutonomous DrivingDomain AdaptationData Augmentation

Positive-Unlabeled Data Purification in the Wild for Object Detection

2021-06-19 · CVPR 2021 1 · Jianyuan Guo, Kai Han, Han Wu, Chao Zhang 외

Deep learning based object detection approaches have achieved great progress with the benefit from large amount of labeled images. However, image annotation remains a laborious, time-consuming and error-prone process…

Knowledge Distillationobject-detectionObject Detection