paper-with-me

Object Localization

18개 벤치마크 · 논문 724편 · 이 태스크의 논문 보기 →

Benchmarks

REVERIE

결과 13개

IllusionVQA

결과 9개

GRIT

결과 3개

KITTI Cars Easy

결과 2개

KITTI Cars Hard

결과 2개

KITTI Cars Moderate

결과 2개

KITTI Cyclists Easy

결과 2개

KITTI Cyclists Hard

결과 2개

KITTI Pedestrian Easy

결과 1개

Mall

결과 1개

PASCAL VOC 2007

결과 1개

PASCAL VOC 2012

결과 1개

Plant

결과 1개

Pupil

결과 1개

Most implemented

Mask R-CNN

2017-03-20 · 구현 179개

Microsoft COCO: Common Objects in Context

2014-05-01 · 구현 38개

Papers

Tail-Likelihood Reinforcement Learning

2026-09-02 · Shrinivas Ramasubramanian, Daman Arora, Fahim Tajwar, Guanning Zeng 외 arxiv

Reinforcement learning typically optimizes average reward. For generative policies, the average can hide an important distinction: two policies can achieve the same mean reward while having very different chances of prod…

Reinforcement LearningObject Localization

WALDO: One-Shot Exemplar-Conditioned Object Detection in Cluttered Scenes

2026-08-28 · Kishor Datta Gupta, Ahmed Rafi Hasan, Md. Mahfuzur Rahman, Md. Sadman Haque 외 arxiv

Locating a specific object instance in a cluttered scene using a single reference image and a short description, and reporting when that instance is absent, large vision-language models usually address this task. We ask …

Object LocalizationObject Detection

Grounding Isn't Knowing: Do VLMs Need Object Localization for Spatial Reasoning?

2026-08-24 · Xiwei Liu, Yulong Li, Xinlin Zhuang, Xuhui Li 외 arxiv

Vision-language models (VLMs) can answer spatial questions, yet the mechanisms connecting object grounding to spatial reasoning remain poorly understood. It is underexplored whether spatial reasoning internally requires …

Object LocalizationSpatial Reasoning

Doomed to Re-Annotate, Forever: The ImageNet Story

2026-08-13 · Illia Volkov, Nikita Kisel, Tetiana Mishkina, Klara Janouskova 외 arxiv

Top-1 accuracy on ImageNet-1k remains the most commonly reported metric in visual recognition. Quality issues with the dataset have been repeatedly reported, yet the original 2012 noisy labels are still predominantly use…

Object Localization

DRPFNet: Dual-domain Residual Progressive Fusion Network for RGB-Thermal Object Detection

2026-08-04 · Zian Wang, Changchun Li arxiv

RGB-thermal (RGB-T) object detection aims to fuse complementary information from visible and thermal modalities to achieve robust detection under varying illumination and weather conditions. Current methods typically emp…

Object LocalizationObject Detection

RadYOLO: Computationally Efficient 3D Object Detection and Segmentation in CT and MRI

2026-08-01 · Kai Geissler, Laurens Müller-Groh, Hans Meine arxiv

Object detection and segmentation in three-dimensional medical images is a very active area of research. However, most proposed deep learning models carry a high computational cost, and only few aim to be broadly applica…

3D Object DetectionObject Localization

전체 724편 보기 →