BigDetection: A Large-scale Benchmark for Improved Object Detector Pre-training
Multiple datasets and open challenges for object detection have been introduced in recent years. To build more general and powerful object detection systems, in this paper, we construct a new large-scale benchmark termed BigDetection. Our goal is to simply leverage the training data from existing datasets (LVIS, OpenImages and Object365) with carefully designed principles, and curate a larger dataset for improved detector pre-training. Specifically, we generate a new taxonomy which unifies the heterogeneous label spaces from different sources. Our BigDetection dataset has 600 object categories and contains over 3.4M training images with 36M bounding boxes. It is much larger in multiple dimensions than previous benchmarks, which offers both opportunities and challenges. Extensive experiments demonstrate its validity as a new benchmark for evaluating different object detection methods, and its effectiveness as a pre-training dataset.
Code (2)
Tasks
Objectobject-detectionObject DetectionSimilar Papers 제목 키워드 기반
Scale-Robust Localization Using General Object Landmarks
Visual localization under large changes in scale is an important capability in many robotic mapping applications, such as localizing at low altitudes in maps built at high altitudes, or performing loop closure over long …
ObjectVisual LocalizationDigital Twin Catalog: A Large-Scale Photorealistic 3D Object Digital Twin Dataset
We introduce the Digital Twin Catalog (DTC), a new large-scale photorealistic 3D object digital twin dataset. A digital twin of a 3D object is a highly detailed, virtually indistinguishable representation of a physical o…
3D Object Reconstruction3D ReconstructionInverse RenderingObject+1Is Pseudo-Lidar needed for Monocular 3D Object detection?
Recent progress in 3D object detection from single images leverages monocular depth estimation as a way to produce 3D pointclouds, turning cameras into pseudo-lidar sensors. These two-stage detectors improve with the acc…
3D Object DetectionDepth EstimationMonocular 3D Object DetectionMonocular Depth Estimation+3Scale Optimization for Full-Image-CNN Vehicle Detection
Many state-of-the-art general object detection methods make use of shared full-image convolutional features (as in Faster R-CNN). This achieves a reasonable test-phase computation time while enjoys the discriminative pow…
Objectobject-detectionObject DetectionRegion Proposal+1Implicit-Scale 3D Reconstruction for Multi-Food Volume Estimation from Monocular Images
We present Implicit-Scale 3D Reconstruction from Monocular Multi-Food Images, a benchmark dataset designed to advance geometry-based food portion estimation in realistic dining scenarios. Existing dietary assessment meth…
3D Reconstruction