paper-with-me

홈 › Papers

Beyond Model Size: Probing the Gaps in Visual in-Context Learning by Training a Tiny Model

2026-06-09 · Sunil Khatri, Steven Landgraf, Markus Ulrich, Simon Reiß arxiv

Visual in-Context Learning (VICL) aims at making progress towards adaptive vision models, that can -- based on a few examples -- adapt to a new task at test-time. With the history of in-context learning in natural language processing research, where large, parameter-heavy models are in use, one pathway that current VICL methods take is model- and data-scaling as key ingredients. Yet, it is not clear, whether these ingredients are the key for in-context learning to take shape in vision models. To stress-test such large models, we challenge them with an extreme counterexample: we train a tiny visual in-context model with merely $1$ million parameters and a modest amount of $70,000$ images. We compare the results of this severely capacity capped tiny model to $7,000\times$ larger VICL models in different adaptive settings, (1) on image data with small distribution shifts, (2) on unseen task encodings and (3) on a completely new task, i.e., the setting VICL envisions. With the chasm of training resources between the tiny- and large models, our experiments showcase a lack in how adaptive capabilities are measured, with respect to how tasks are encoded, which tasks were used in pre-training and the choice of metrics. These gaps in current VICL benchmarking underscore a need for innovation in evaluation of adaptive capabilities.

📄 PDF Abstract BibTeX arXiv:2606.10905

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Make Your LLM Fully Utilize the Context

2024-04-25 · Shengnan An, Zexiong Ma, Zeqi Lin, Nanning Zheng 외

While many contemporary large language models (LLMs) can process lengthy input, they still struggle to fully utilize information within the long context, known as the lost-in-the-middle challenge. We hypothesize that it …

4kInformation RetrievalMMLURetrieval

Visual Probing: Cognitive Framework for Explaining Self-Supervised Image Representations

2021-06-21 · Witold Oleszkiewicz, Dominika Basaj, Igor Sieradzki, Michał Górszczak 외

Recently introduced self-supervised methods for image representation learning provide on par or superior results to their fully supervised competitors, yet the corresponding efforts to explain the self-supervised approac…

Representation Learning

Do We Know What LLMs Don't Know? A Study of Consistency in Knowledge Probing

2025-05-27 · Raoyuan Zhao, Abdullatif Köksal, Ali Modarressi, Michael A. Hedderich 외

The reliability of large language models (LLMs) is greatly compromised by their tendency to hallucinate, underscoring the need for precise identification of knowledge gaps within LLMs. Various methods for probing such ga…

Knowledge Probing

Probing Brain Context-Sensitivity with Masked-Attention Generation

2023-05-23 · Alexandre Pasquiou, Yair Lakretz, Bertrand Thirion, Christophe Pallier

Two fundamental questions in neurolinguistics concerns the brain regions that integrate information beyond the lexical level, and the size of their window of integration. To address these questions we introduce a new app…

SensitivityWord Embeddings

Black-box language model explanation by context length probing

2022-12-30 · Ondřej Cífka, Antoine Liutkus

The increasingly widespread adoption of large language models has highlighted the need for improving their explainability. We present context length probing, a novel explanation technique for causal language models, base…

Language ModelingLanguage Modelling