paper-with-me

홈 › Papers

BigDetection: A Large-scale Benchmark for Improved Object Detector Pre-training

2022-03-24 · Likun Cai, Zhi Zhang, Yi Zhu, Li Zhang, Mu Li, xiangyang xue

Multiple datasets and open challenges for object detection have been introduced in recent years. To build more general and powerful object detection systems, in this paper, we construct a new large-scale benchmark termed BigDetection. Our goal is to simply leverage the training data from existing datasets (LVIS, OpenImages and Object365) with carefully designed principles, and curate a larger dataset for improved detector pre-training. Specifically, we generate a new taxonomy which unifies the heterogeneous label spaces from different sources. Our BigDetection dataset has 600 object categories and contains over 3.4M training images with 36M bounding boxes. It is much larger in multiple dimensions than previous benchmarks, which offers both opportunities and challenges. Extensive experiments demonstrate its validity as a new benchmark for evaluating different object detection methods, and its effectiveness as a pre-training dataset.

📄 PDF Abstract BibTeX arXiv:2203.13249

Code (2)

amazon-research/bigdetection 공식 구현 pytorch
amazon-science/bigdetection pytorch

Tasks

Objectobject-detectionObject Detection

Similar Papers 제목 키워드 기반

Scale-Robust Localization Using General Object Landmarks

2017-10-28 · Andrew Holliday, Gregory Dudek

Visual localization under large changes in scale is an important capability in many robotic mapping applications, such as localizing at low altitudes in maps built at high altitudes, or performing loop closure over long …

ObjectVisual Localization

Digital Twin Catalog: A Large-Scale Photorealistic 3D Object Digital Twin Dataset

2025-04-11 · CVPR 2025 1 · Zhao Dong, Ka Chen, Zhaoyang Lv, Hong-Xing Yu 외

We introduce the Digital Twin Catalog (DTC), a new large-scale photorealistic 3D object digital twin dataset. A digital twin of a 3D object is a highly detailed, virtually indistinguishable representation of a physical o…

3D Object Reconstruction3D ReconstructionInverse RenderingObject+1

Is Pseudo-Lidar needed for Monocular 3D Object detection?

2021-08-13 · ICCV 2021 10 · Dennis Park, Rares Ambrus, Vitor Guizilini, Jie Li 외

Recent progress in 3D object detection from single images leverages monocular depth estimation as a way to produce 3D pointclouds, turning cameras into pseudo-lidar sensors. These two-stage detectors improve with the acc…

3D Object DetectionDepth EstimationMonocular 3D Object DetectionMonocular Depth Estimation+3

Scale Optimization for Full-Image-CNN Vehicle Detection

2018-02-20 · Yang Gao, Shouyan Guo, Kaimin Huang, Jiaxin Chen 외

Many state-of-the-art general object detection methods make use of shared full-image convolutional features (as in Faster R-CNN). This achieves a reasonable test-phase computation time while enjoys the discriminative pow…

Objectobject-detectionObject DetectionRegion Proposal+1

Implicit-Scale 3D Reconstruction for Multi-Food Volume Estimation from Monocular Images

2026-02-13 · Yuhao Chen, Gautham Vinod, Siddeshwar Raghavan, Talha Ibn Mahmud 외 arxiv

We present Implicit-Scale 3D Reconstruction from Monocular Multi-Food Images, a benchmark dataset designed to advance geometry-based food portion estimation in realistic dining scenarios. Existing dietary assessment meth…

3D Reconstruction