paper-with-me

홈 › Papers

Probing the Decision Boundaries of In-context Learning in Large Language Models

2024-06-17 · Siyan Zhao, Tung Nguyen, Aditya Grover

In-context learning is a key paradigm in large language models (LLMs) that enables them to generalize to new tasks and domains by simply prompting these models with a few exemplars without explicit parameter updates. Many attempts have been made to understand in-context learning in LLMs as a function of model scale, pretraining data, and other factors. In this work, we propose a new mechanism to probe and understand in-context learning from the lens of decision boundaries for in-context binary classification. Decision boundaries are straightforward to visualize and provide important information about the qualitative behavior of the inductive biases of standard classifiers. To our surprise, we find that the decision boundaries learned by current LLMs in simple binary classification tasks are often irregular and non-smooth, regardless of linear separability in the underlying task. This paper investigates the factors influencing these decision boundaries and explores methods to enhance their generalizability. We assess various approaches, including training-free and fine-tuning methods for LLMs, the impact of model architecture, and the effectiveness of active prompting techniques for smoothing decision boundaries in a data-efficient manner. Our findings provide a deeper understanding of in-context learning dynamics and offer practical improvements for enhancing robustness and generalizability of in-context learning.

📄 PDF Abstract BibTeX arXiv:2406.11233

Code (1)

siyan-zhao/ICL_decision_boundary 공식 구현 pytorch

Tasks

Binary ClassificationIn-Context Learning

Similar Papers 제목 키워드 기반

Bias Beyond Demographics: Probing Decision Boundaries in Black-Box LVLMs via Counterfactual VQA

2025-08-05 · Zaiying Zhao, Toshihiko Yamasaki arxiv

Recent advances in large vision-language models (LVLMs) have amplified concerns about fairness, yet existing evaluations remain confined to demographic attributes and often conflate fairness with refusal behavior. This p…

Multimodal Reasoning

Light or Full Verb? A Minimal-Pair Dataset for Probing Phraseological Competence in Language Models

2026-06-03 · Francesca Franzon, Nicolas Rosàs Gómez, Leo Wanner arxiv

Frequent verbs such as 'have' and 'make' can function either as collocates in light-verb constructions or as full lexical predicates, as in 'make a decision' vs. 'make a cake'. Whether language models represent this dist…

Delineating Knowledge Boundaries for Honest Large Vision-Language Models

2026-04-29 · Junru Song, Yimeng Hu, Yijing Chen, Huining Li 외 arxiv

Large Vision-Language Models (VLMs) have achieved remarkable multimodal performance yet remain prone to factual hallucinations, particularly in long-tail or specialized domains. Moreover, current models exhibit a weak ca…

BabyLM's First Words: Word Segmentation as a Phonological Probing Task

2025-04-04 · Zébulon Goriely

Language models provide a key framework for studying linguistic theories based on prediction, but phonological analysis using large language models (LLMs) is difficult; there are few phonological benchmarks beyond Englis…

Black-box language model explanation by context length probing

2022-12-30 · Ondřej Cífka, Antoine Liutkus

The increasingly widespread adoption of large language models has highlighted the need for improving their explainability. We present context length probing, a novel explanation technique for causal language models, base…

Language ModelingLanguage Modelling