paper-with-me

홈 › Papers

Toward Robust In-Context Learning: Leveraging Out-of-distribution Proxies for Target Inaccessible Demonstration Retrieval

2026-04-13 · Hao Xu, Rite Bo, Fausto Giunchiglia, Yingji Li, Rui Song arxiv

Although studies have demonstrated that Large Language Models (LLMs) can perform well on Out-of-Distribution (OOD) tasks, their advantage tends to diminish as the distribution shift becomes more severe. Consequently, researchers aim to retrieve distributionally similar and informative demonstrations from the available source domain to boost the inference capabilities of LLMs. However, in practical scenarios where the target domain is inaccessible, evaluating the unknown distribution is challenging, which indirectly impacts the quality of the selected demonstrations. To address this problem, we propose \textbf{DOPA}, a demonstration search framework that incorporates an OOD proxy to approximate the inaccessible target domain and guide the retrieval process. Building on proxy-based evaluation, DOPA further introduces a Mahalanobis distance-based global diversity constraint to ensure sufficient diversity among the retrieved demonstrations. Experimental results on multiple LLMs and tasks demonstrate that DOPA effectively enhances robustness in OOD settings\footnote{https://github.com/bort64/ood\_code}.

📄 PDF Abstract BibTeX arXiv:2606.00014

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Weak Proxies are Sufficient and Preferable for Fairness with Missing Sensitive Attributes

2022-10-06 · Zhaowei Zhu, Yuanshun Yao, Jiankai Sun, Hang Li 외

Evaluating fairness can be challenging in practice because the sensitive attributes of data are often inaccessible due to privacy constraints. The go-to approach that the industry frequently adopts is using off-the-shelf…

Fairness

MetaMoE: Diversity-Aware Proxy Selection for Privacy-Preserving Mixture-of-Experts Unification

2026-05-14 · Weisen Jiang, Shuhao Chen, Sinno Jialin Pan arxiv

Mixture-of-Experts (MoE) models scale capacity by combining specialized experts, but most existing approaches assume centralized access to training data. In practice, data are distributed across clients and cannot be sha…

MLDGG: Meta-Learning for Domain Generalization on Graphs

2024-11-19 · Qin Tian, Chen Zhao, Minglai Shao, Wenjun Wang 외

Domain generalization on graphs aims to develop models with robust generalization capabilities, ensuring effective performance on the testing set despite disparities between testing and training distributions. However, e…

Domain GeneralizationMeta-LearningTransfer Learning

Large Language Models as Proxies for Theories of Human Linguistic Cognition

2025-02-11 · Imry Ziv, Nur Lan, Emmanuel Chemla, Roni Katzir

We consider the possible role of current large language models (LLMs) in the study of human linguistic cognition. We focus on the use of such models as proxies for theories of cognition that are relatively linguistically…

Causal Information Splitting: Engineering Proxy Features for Robustness to Distribution Shifts

2023-05-10 · Bijan Mazaheri, Atalanti Mastakouri, Dominik Janzing, Michaela Hardt

Statistical prediction models are often trained on data from different probability distributions than their eventual use cases. One approach to proactively prepare for these shifts harnesses the intuition that causal mec…

counterfactualfeature selection