paper-with-me

홈 › Papers

From Pre-training Corpora to Large Language Models: What Factors Influence LLM Performance in Causal Discovery Tasks?

2024-07-29 · Tao Feng, Lizhen Qu, Niket Tandon, Zhuang Li, Xiaoxi Kang, Gholamreza Haffari

Recent advances in artificial intelligence have seen Large Language Models (LLMs) demonstrate notable proficiency in causal discovery tasks. This study explores the factors influencing the performance of LLMs in causal discovery tasks. Utilizing open-source LLMs, we examine how the frequency of causal relations within their pre-training corpora affects their ability to accurately respond to causal discovery queries. Our findings reveal that a higher frequency of causal mentions correlates with better model performance, suggesting that extensive exposure to causal information during training enhances the models' causal discovery capabilities. Additionally, we investigate the impact of context on the validity of causal relations. Our results indicate that LLMs might exhibit divergent predictions for identical causal relations when presented in different contexts. This paper provides the first comprehensive analysis of how different factors contribute to LLM performance in causal discovery tasks.

📄 PDF Abstract BibTeX arXiv:2407.19638

Code (0)

등록된 구현이 없습니다.

Tasks

Causal Discovery

Similar Papers 제목 키워드 기반

Extracting Linguistic Knowledge from Speech: A Study of Stop Realization in 5 Romance Languages

2022-06-01 · LREC 2022 6 · Yaru Wu, Mathilde Hutin, Ioana Vasilescu, Lori Lamel 외

This paper builds upon recent work in leveraging the corpora and tools originally used to develop speech technologies for corpus-based linguistic studies. We address the non-canonical realization of consonants in connect…

speech-recognitionSpeech Recognition

What's In My Big Data?

2023-10-31 · Yanai Elazar, Akshita Bhagia, Ian Magnusson, Abhilasha Ravichander 외

Large text corpora are the backbone of language models. However, we have a limited understanding of the content of these corpora, including general statistics, quality, social factors, and inclusion of evaluation data (c…

Benchmarking

Large Human Language Models: A Need and the Challenges

2023-11-09 · Nikita Soni, H. Andrew Schwartz, João Sedoc, Niranjan Balasubramanian

As research in human-centered NLP advances, there is a growing recognition of the importance of incorporating human and social factors into NLP models. At the same time, our NLP systems have become heavily reliant on LLM…

Inducing Discourse Marker Inventories from Lexical Knowledge Graphs

2022-06-01 · LREC 2022 6 · Christian Chiarcos

Discourse marker inventories are important tools for the development of both discourse parsers and corpora with discourse annotations. In this paper we explore the potential of massively multilingual lexical knowledge gr…

Knowledge Graphs

Synthetic Pre-Training Tasks for Neural Machine Translation

2022-12-19 · Zexue He, Graeme Blackwood, Rameswar Panda, Julian McAuley 외

Pre-training models with large crawled corpora can lead to issues such as toxicity and bias, as well as copyright and privacy concerns. A promising way of alleviating such concerns is to conduct pre-training with synthet…

Machine TranslationNMTTranslation