paper-with-me

홈 › Papers

Humans overrely on overconfident language models, across languages

2025-07-08 · Neil Rathi, Dan Jurafsky, Kaitlyn Zhou arxiv

As large language models (LLMs) are deployed globally, it is crucial that their responses are calibrated across languages to accurately convey uncertainty and limitations. Prior work shows that LLMs are linguistically overconfident in English, leading users to overrely on confident generations. However, the usage and interpretation of epistemic markers (e.g., 'I think it's') differs sharply across languages. Here, we study the risks of multilingual linguistic (mis)calibration, overconfidence, and overreliance across five languages to evaluate LLM safety in a global context. Our work finds that overreliance risks are high across languages. We first analyze the distribution of LLM-generated epistemic markers and observe that LLMs are overconfident across languages, frequently generating strengtheners even as part of incorrect responses. Model generations are, however, sensitive to documented cross-linguistic variation in usage: for example, models generate the most markers of uncertainty in Japanese and the most markers of certainty in German and Mandarin. Next, we measure human reliance rates across languages, finding that reliance behaviors differ cross-linguistically: for example, participants are significantly more likely to discount expressions of uncertainty in Japanese than in English (i.e., ignore their 'hedging' function and rely on generations that contain them). Taken together, these results indicate a high risk of reliance on overconfident model generations across languages. Our findings highlight the challenges of multilingual linguistic calibration and stress the importance of culturally and linguistically contextualized model safety evaluations.

📄 PDF Abstract BibTeX arXiv:2507.06306

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Translate-R1: Cost-Aware Translation Tool Use via Reinforcement Learning

2026-06-05 · Pratik Jayarao, Chaitanya Dwivedi, Himanshu Gupta, Neeraj Varshney 외 arxiv

The performance gap across languages in LLMs is well documented, and closing it natively requires pretraining or fine-tuning on corpora that, for most languages, do not exist. Translation offers an alternative: convertin…

Reinforcement Learning

The Riddle of Reflection: Evaluating Reasoning and Self-Awareness in Multilingual LLMs using Indian Riddles

2025-11-02 · Abhinav P M, Ojasva Saxena, Oswald C, Parameswari Krishnamurthy arxiv

The extent to which large language models (LLMs) can perform culturally grounded reasoning across non-English languages remains underexplored. This paper examines the reasoning and self-assessment abilities of LLMs acros…

ChatGPT Prompting Cannot Estimate Predictive Uncertainty in High-Resource Languages

2023-11-10 · Martino Pelucchi, Matias Valdenegro-Toro

ChatGPT took the world by storm for its impressive abilities. Due to its release without documentation, scientists immediately attempted to identify its limits, mainly through its performance in natural language processi…

Comparing Styles across Languages: A Cross-Cultural Exploration of Politeness

2023-10-11 · Shreya Havaldar, Matthew Pressimone, Eric Wong, Lyle Ungar

Understanding how styles differ across languages is advantageous for training both humans and computers to generate culturally appropriate text. We introduce an explanation framework to extract stylistic differences from…

Tracing the ongoing emergence of human-like reasoning in Large Language Models

2026-05-20 · Paolo Morosi, Nikoleta Pantelidou, Fritz Günther, Elena Pagliarini 외 arxiv

Humans effortlessly go beyond literal meanings: If you mow the lawn, I will give you fifty dollars, is typically understood as implying that the speaker will pay only if the lawn is mowed, whereas If you are hungry, ther…

Logical Reasoning