paper-with-me

홈 › Papers

Exploring Language Model Generalization in Low-Resource Extractive QA

2024-09-27 · Saptarshi Sengupta, Wenpeng Yin, Preslav Nakov, Shreya Ghosh, Suhang Wang

In this paper, we investigate Extractive Question Answering (EQA) with Large Language Models (LLMs) under domain drift, i.e., can LLMs generalize to domains that require specific knowledge such as medicine and law in a zero-shot fashion without additional in-domain training? To this end, we devise a series of experiments to explain the performance gap empirically. Our findings suggest that: (a) LLMs struggle with dataset demands of closed domains such as retrieving long answer spans; (b) Certain LLMs, despite showing strong overall performance, display weaknesses in meeting basic requirements as discriminating between domain-specific senses of words which we link to pre-processing decisions; (c) Scaling model parameters is not always effective for cross domain generalization; and (d) Closed-domain datasets are quantitatively much different than open-domain EQA datasets and current LLMs struggle to deal with them. Our findings point out important directions for improving existing LLMs.

📄 PDF Abstract BibTeX arXiv:2409.18446

Code (1)

saptarshi059/generalization-hypothesis 공식 구현 pytorch

Tasks

Domain GeneralizationExtractive Question-AnsweringLanguage ModelingLanguage ModellingQuestion Answering

Similar Papers 제목 키워드 기반

Extractive Fact Decomposition for Interpretable Natural Language Inference in one Forward Pass

2025-09-23 · Nicholas Popovič, Michael Färber arxiv

Recent works in Natural Language Inference (NLI) and related tasks, such as automated fact-checking, employ atomic fact decomposition to enhance interpretability and robustness. For this, existing methods rely on resourc…

Natural Language Inference

Extractive Structures Learned in Pretraining Enable Generalization on Finetuned Facts

2024-12-05 · Jiahai Feng, Stuart Russell, Jacob Steinhardt

Pretrained language models (LMs) can generalize to implications of facts that they are finetuned on. For example, if finetuned on ``John Doe lives in Tokyo," LMs can correctly answer ``What language do the people in John…

counterfactual

Exploring Content Selection in Summarization of Novel Chapters

2020-05-04 · ACL 2020 6 · Faisal Ladhak, Bryan Li, Yaser Al-Onaizan, Kathleen McKeown

We present a new summarization task, generating summaries of novel chapters using summary/chapter pairs from online study guides. This is a harder task than the news summarization task, given the chapter length as well a…

Extractive SummarizationNews Summarization

Extractive and Abstractive Explanations for Fact-Checking and Evaluation of News

2021-04-27 · NAACL (NLP4IF) 2021 6 · Ashkan Kazemi, Zehua Li, Verónica Pérez-Rosas, Rada Mihalcea

In this paper, we explore the construction of natural language explanations for news claims, with the goal of assisting fact-checking and news evaluation applications. We experiment with two methods: (1) an extractive me…

Fact CheckingLanguage ModelingLanguage ModellingMisinformation

Exploring Multitask Learning for Low-Resource AbstractiveSummarization

2021-09-17 · Ahmed Magooda, Mohamed Elaraby, Diane Litman

This paper explores the effect of using multitask learning for abstractive summarization in the context of small training corpora. In particular, we incorporate four different tasks (extractive summarization, language mo…

Abstractive Text SummarizationExtractive SummarizationLanguage ModelingLanguage Modelling