paper-with-me

Papers

Instance-Level Generation for Representation Learning

2025-10-10 · Yankun Wu, Zakaria Laskar, Giorgos Kordopatis-Zilos, Noa Garcia, Giorgos Tolias arxiv

Instance-level recognition (ILR) focuses on identifying individual objects rather than broad categories, offering the highest granularity in image classification. However, this fine-grained nature makes creating large-scale annotated datasets challenging, limiting ILR's real-world applicability across domains. To overcome this, we introduce a novel approach that synthetically generates diverse object instances from multiple domains under varied conditions and backgrounds, forming a large-scale training set. Unlike prior work on automatic data synthesis, our method is the first to address ILR-specific challenges without relying on any real images. Fine-tuning foundation vision models on the generated data significantly improves retrieval performance across seven ILR benchmarks spanning multiple domains. Our approach offers a new, efficient, and effective alternative to extensive data collection and curation, introducing a new ILR paradigm where the only input is the names of the target domains, unlocking a wide range of real-world applications.

📄 PDF Abstract BibTeX arXiv:2510.09171

Code (0)

등록된 구현이 없습니다.

Tasks

Representation LearningImage Classification

Similar Papers 제목 키워드 기반

Improving Discriminative Visual Representation Learning via Automatic Mixup

2021-09-29 · Siyuan Li, Zicheng Liu, Di wu, Stan Z. Li

Mixup, a convex interpolation technique for data augmentation, has achieved great success in deep neural networks. However, the community usually confines it to supervised scenarios or applies it as a predefined augmenta…

Data AugmentationRepresentation Learning

CC-FMO: Camera-Conditioned Zero-Shot Single Image to 3D Scene Generation with Foundation Model Orchestration

2025-11-29 · Boshi Tang, Henry Zheng, Rui Huang, Gao Huang arxiv

High-quality 3D scene generation from a single image is crucial for AR/VR and embodied AI applications. Early approaches struggle to generalize due to reliance on specialized models trained on curated small datasets. Whi…

Scene GenerationPose Estimation

Clustering-based Tile Embedding (CTE): A General Representation for Level Design with Skewed Tile Distributions

2022-10-23 · Mrunal Jadhav, Matthew Guzdial

There has been significant research interest in Procedural Level Generation via Machine Learning (PLGML), applying ML techniques to automated level generation. One recent trend is in the direction of learning representat…

Clustering

Keywords and Instances: A Hierarchical Contrastive Learning Framework Unifying Hybrid Granularities for Text Generation

2022-05-26 · ACL 2022 5 · Mingzhe Li, Xiexiong Lin, Xiuying Chen, Jinxiong Chang 외

Contrastive learning has achieved impressive success in generation tasks to militate the "exposure bias" problem and discriminatively exploit the different quality of references. Existing works mostly focus on contrastiv…

Contrastive LearningDialogue GenerationSentenceText Generation

InstanceV: Instance-Level Video Generation

2025-11-28 · Yuheng Chen, Teng Hu, Jiangning Zhang, Zhucun Xue 외 arxiv

Recent advances in text-to-video diffusion models have enabled the generation of high-quality videos conditioned on textual descriptions. However, most existing text-to-video models rely solely on textual conditions, lac…

Video Generation