paper-with-me

홈 › Papers

More than Correlation: Do Large Language Models Learn Causal Representations of Space?

2023-12-26 · Yida Chen, Yixian Gan, Sijia Li, Li Yao, Xiaohan Zhao

Recent work found high mutual information between the learned representations of large language models (LLMs) and the geospatial property of its input, hinting an emergent internal model of space. However, whether this internal space model has any causal effects on the LLMs' behaviors was not answered by that work, led to criticism of these findings as mere statistical correlation. Our study focused on uncovering the causality of the spatial representations in LLMs. In particular, we discovered the potential spatial representations in DeBERTa, GPT-Neo using representational similarity analysis and linear and non-linear probing. Our casual intervention experiments showed that the spatial representations influenced the model's performance on next word prediction and a downstream task that relies on geospatial information. Our experiments suggested that the LLMs learn and use an internal model of space in solving geospatial related tasks.

📄 PDF Abstract BibTeX arXiv:2312.16257

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

How do I file a dispute with Expedia?*DisputeFastService How do I file a dispute with Expedia? To file a dispute with Expedia, call +1(888) (829) (0881) OR +1(805) (330) (4056), or use their Help Center to submit your case with…
DeBERTa DeBERTa is a Transformer-based neural language model that aims to improve the…
GPT-Neo An implementation of model & data parallel GPT3-like models using the mesh-tensorflow…

Similar Papers 제목 키워드 기반

Gender Coreference and Bias Evaluation at WMT 2020

2020-10-12 · WMT (EMNLP) 2020 11 · Tom Kocmi, Tomasz Limisiewicz, Gabriel Stanovsky

Gender bias in machine translation can manifest when choosing gender inflections based on spurious gender correlations. For example, always translating doctors as men and nurses as women. This can be particularly harmful…

Machine TranslationTranslation

Correlation Dimension of Natural Language in a Statistical Manifold

2024-05-10 · Xin Du, Kumiko Tanaka-Ishii

The correlation dimension of natural language is measured by applying the Grassberger-Procaccia algorithm to high-dimensional sequences produced by a large-scale language model. This method, previously studied only in a …

Language ModelingLanguage Modelling

Market correlation structure changes around the Great Crash

2016-01-30

We perform a comparative analysis of the Chinese stock market around the occurrence of the 2008 crisis based on the random matrix analysis of high-frequency stock returns of 1228 stocks listed on the Shanghai and Shenzhe…

Could We Have Had Better Multilingual LLMs If English Was Not the Central Language?

2024-02-21 · Ryandito Diandaru, Lucky Susanto, Zilu Tang, Ayu Purwarianti 외

Large Language Models (LLMs) demonstrate strong machine translation capabilities on languages they are trained on. However, the impact of factors beyond training data size on translation performance remains a topic of de…

Machine TranslationTranslation

Word Familiarity and Frequency

2018-06-09 · Kumiko Tanaka-Ishii, Hiroshi Terada

Word frequency is assumed to correlate with word familiarity, but the strength of this correlation has not been thoroughly investigated. In this paper, we report on our analysis of the correlation between a word familiar…