paper-with-me

Papers

3D Copy-Paste: Physically Plausible Object Insertion for Monocular 3D Detection

2023-12-08 · NeurIPS 2023 11 · Yunhao Ge, Hong-Xing Yu, Cheng Zhao, Yuliang Guo, Xinyu Huang, Liu Ren, Laurent Itti, Jiajun Wu

A major challenge in monocular 3D object detection is the limited diversity and quantity of objects in real datasets. While augmenting real scenes with virtual objects holds promise to improve both the diversity and quantity of the objects, it remains elusive due to the lack of an effective 3D object insertion method in complex real captured scenes. In this work, we study augmenting complex real indoor scenes with virtual objects for monocular 3D object detection. The main challenge is to automatically identify plausible physical properties for virtual assets (e.g., locations, appearances, sizes, etc.) in cluttered real scenes. To address this challenge, we propose a physically plausible indoor 3D object insertion approach to automatically copy virtual objects and paste them into real scenes. The resulting objects in scenes have 3D bounding boxes with plausible physical locations and appearances. In particular, our method first identifies physically feasible locations and poses for the inserted objects to prevent collisions with the existing room layout. Subsequently, it estimates spatially-varying illumination for the insertion location, enabling the immersive blending of the virtual objects into the original scene with plausible appearances and cast shadows. We show that our augmentation method significantly improves existing monocular 3D object models and achieves state-of-the-art performance. For the first time, we demonstrate that a physically plausible 3D object insertion, serving as a generative data augmentation technique, can lead to significant improvements for discriminative downstream tasks such as monocular 3D object detection. Project website: https://gyhandy.github.io/3D-Copy-Paste/

📄 PDF Abstract BibTeX arXiv:2312.05277

Code (1)

gyhandy/3d-copy-paste 공식 구현

Tasks

3D Object DetectionData AugmentationDiversityMonocular 3D Object DetectionObjectobject-detectionObject Detection

Similar Papers 제목 키워드 기반

Depth-Copy-Paste: Multimodal and Depth-Aware Compositing for Robust Face Detection

2025-12-12 · Qiushi Guo arxiv

Data augmentation is crucial for improving the robustness of face detection systems, especially under challenging conditions such as occlusion, illumination variation, and complex environments. Traditional copy paste aug…

Data AugmentationFace Detection

X-Paste: Revisiting Scalable Copy-Paste for Instance Segmentation using CLIP and StableDiffusion

2022-12-07 · Hanqing Zhao, Dianmo Sheng, Jianmin Bao, Dongdong Chen 외

Copy-Paste is a simple and effective data augmentation strategy for instance segmentation. By randomly pasting object instances onto new background images, it creates new training data for free and significantly boosts t…

Data AugmentationInstance SegmentationObjectObject Detection+4

Simple Copy-Paste is a Strong Data Augmentation Method for Instance Segmentation

2020-12-13 · CVPR 2021 1 · Golnaz Ghiasi, Yin Cui, Aravind Srinivas, Rui Qian 외

Building instance segmentation models that are data-efficient and can handle rare object categories is an important challenge in computer vision. Leveraging data augmentations is a promising direction towards addressing …

Data AugmentationImage AugmentationInstance SegmentationObject Detection+2

Copy-Trasform-Paste: Zero-Shot Object-Object Alignment Guided by Vision-Language and Geometric Constraints

2026-01-20 · Rotem Gatenyo, Ohad Fried arxiv

We study zero-shot 3D alignment of two given meshes, using a text prompt describing their spatial relation -- an essential capability for content creation and scene assembly. Earlier approaches primarily rely on geometri…

SDI-Paste: Synthetic Dynamic Instance Copy-Paste for Video Instance Segmentation

2024-10-16 · Sahir Shrestha, Weihao Li, Gao Zhu, Nick Barnes

Data augmentation methods such as Copy-Paste have been studied as effective ways to expand training datasets while incurring minimal costs. While such methods have been extensively implemented for image level tasks, we f…

Data AugmentationInstance SegmentationMORPHSemantic Segmentation+1