paper-with-me

Papers

Object-Centric Dataset Resources for Constrained-Data Image Generation and Augmentation

2026-06-19 · Vasile Marian, Yong-Bin Kang, Alexander Buddery arxiv

Object-centric image generation is important in settings with few labeled examples, including pedestrian analysis in smart-city scenes, traffic-sign inspection, and domain-specific object detection. Synthetic images are most useful for training and evaluation when datasets preserve object structure, bounding boxes, visual diversity, and realistic context. Existing image datasets usually target classification, detection, or scene understanding rather than controlled object-centric generation and augmentation with limited class-specific data. We present a shareable collection of three object-centric dataset resources: Cityscapes-Pedestrian, TrafficSigns, and COCO PottedPlant. The collection standardizes 256-by-256 object-centric crops and bounding-box annotations across three regimes: dense pedestrian scenes with privacy blur and occlusion, cleaner high-contrast traffic signs, and context-diverse potted-plant scenes. The release contains 3,009 TrafficSigns samples, 2,156 Cityscapes-Pedestrian manifest records, and 7,679 COCO PottedPlant manifest records. The larger COCO-derived manifest preserves contextual and multi-instance diversity, while equal-size subsets can be drawn with a fixed random seed for controlled comparisons. The release provides direct TrafficSigns data where redistribution is permitted, together with scripts, manifests, box-level annotation tables, checksums, and reconstruction documentation for the Cityscapes- and COCO-derived subsets. It is available through the Latzi/object-centric-low-data-datasets GitHub repository and Zenodo DOI 10.5281/zenodo.20573001. The collection supports label and split inspection, subset creation, reconstruction from upstream data, and evaluation of object-centric image generation or synthetic-data augmentation methods on shared records.

📄 PDF Abstract BibTeX arXiv:2606.21113

Code (0)

등록된 구현이 없습니다.

Tasks

Scene UnderstandingData AugmentationObject DetectionImage Generation

Similar Papers 제목 키워드 기반

Object-Centric Learning for Real-World Videos by Predicting Temporal Feature Similarities

2023-06-07 · NeurIPS 2023 11 · Andrii Zadaianchuk, Maximilian Seitzer, Georg Martius

Unsupervised video-based object-centric learning is a promising avenue to learn structured representations from large, unlabeled video collections, but previous approaches have only managed to scale to real-world dataset…

ObjectObject Discovery

EgoInteract: Synthetic Egocentric Videos Generation for Interaction Understanding and Anticipation

2026-05-18 · Rosario Leonardi, Francesco Ragusa, Daniele Materia, Alessandro Passanisi 외 arxiv

Collecting large-scale egocentric video datasets with dense spatial and temporal annotations is costly, slow, and often constrained by environmental biases, privacy constraints, and limited coverage of interaction patter…

Active Object DetectionAction SegmentationVideo Generation

Fine-grained Spatiotemporal Grounding on Egocentric Videos

2025-08-01 · Shuo Liang, Yiwu Zhong, Zi-Yuan Hu, Yeyao Tao 외 arxiv

Spatiotemporal video grounding aims to localize target entities in videos based on textual queries. While existing research has made significant progress in exocentric videos, the egocentric setting remains relatively un…

Video Grounding

Object-Centric Data Synthesis for Category-level Object Detection

2025-11-28 · Vikhyat Agarwal, Jiayi Cora Guo, Declan Hoban, Sissi Zhang 외 arxiv

Deep learning approaches to object detection have achieved reliable detection of specific object classes in images. However, extending a model's detection capability to new object classes requires large amounts of annota…

Object Detection

SlotDiffusion: Object-Centric Generative Modeling with Diffusion Models

2023-05-18 · NeurIPS 2023 11 · Ziyi Wu, Jingyu Hu, Wuyue Lu, Igor Gilitschenski 외

Object-centric learning aims to represent visual data with a set of object entities (a.k.a. slots), providing structured representations that enable systematic generalization. Leveraging advanced architectures like Trans…

Image GenerationObjectObject DiscoverySemantic Segmentation+3