paper-with-me

Papers

Doomed to Re-Annotate, Forever: The ImageNet Story

2026-08-13 · Illia Volkov, Nikita Kisel, Tetiana Mishkina, Klara Janouskova, Jiri Matas arxiv

Top-1 accuracy on ImageNet-1k remains the most commonly reported metric in visual recognition. Quality issues with the dataset have been repeatedly reported, yet the original 2012 noisy labels are still predominantly used. The paper presents a comprehensive effort, which goes well beyond prior correction attempts, towards obtaining accurate and complete ImageNet-1k validation set annotations. The result, ReImageNet, includes multilabel correction, object localization, revised class definitions, and semantic attributes (text-recognition, rendition, reflection, crowd, dominant). The reannotation reveals that approximately 12% of the original ImageNet-1k labels are incorrect, 33.3% of images are multilabel and 3.8% contain no object from an ImageNet-1k class. With the new labels, top-1 accuracy increases by up to 1.2% for supervised models and by 5-6% for MLLMs. We argue that annotation at ImageNet scale cannot realistically be completed in one pass, as errors and definitional issues are discovered only through annotating, and we build our pipeline around repeated refinement and error checking. We observed that human and LLM collaboration with appropriate tooling represents the current quality ceiling for annotation at this scale. ImageNet-1k issues propagate into its derivative test sets, indicating that the problem is structural rather than specific to any single benchmark. All annotations, class definitions, guidelines, and analysis code have been publicly released. Project page: https://vrg.fel.cvut.cz/reimagenet Annotations: https://huggingface.co/datasets/vrg-prague/ReImageNet Code: https://github.com/klarajanouskova/ImageNet

📄 PDF Abstract BibTeX arXiv:2608.13783

Code (0)

등록된 구현이 없습니다.

Tasks

Object Localization

Similar Papers 제목 키워드 기반

Phonological Soundscapes in Medieval Poetry

2017-08-01 · WS 2017 8 · Christopher Hench

The oral component of medieval poetry was integral to its performance and reception. Yet many believe that the medieval voice has been forever lost, and any attempts at rediscovering it are doomed to failure due to scrib…

FOREVER: Forgetting Curve-Inspired Memory Replay for Language Model Continual Learning

2026-01-07 · Yujie Feng, Hao Wang, Jian Li, Xu Chu 외 arxiv

Continual learning (CL) for large language models (LLMs) aims to enable sequential knowledge acquisition without catastrophic forgetting. Memory replay methods are widely used for their practicality and effectiveness, bu…

Continual Learning

Robot Grasping and Manipulation: A Prospective

2023-03-14 · Claudio Zito

``A simple handshake would give them away''. This is how Anthony Hopkins' fictional character, Dr Robert Ford, summarises a particular flaw of the 2016 science-fiction \emph{Westworld}'s hosts. In the storyline, Westworl…

Learn the new, keep the old: Extending pretrained models with new anatomy and images

2018-06-01 · Firat Ozdemir, Philipp Fuernstahl, Orcun Goksel

Deep learning has been widely accepted as a promising solution for medical image segmentation, given a sufficiently large representative dataset of images with corresponding annotations. With ever increasing amounts of a…

AnatomyImage SegmentationIncremental LearningMedical Image Segmentation+2

What is known about Vertex Cover Kernelization?

2018-11-23 · Michael R. Fellows, Lars Jaffke, Aliz Izabella Király, Frances A. Rosamond 외

We are pleased to dedicate this survey on kernelization of the Vertex Cover problem, to Professor Juraj Hromkovi\v{c} on the occasion of his 60th birthday. The Vertex Cover problem is often referred to as the Drosophila …

Survey