paper-with-me

홈 › Papers

Uncovering the Fragility of Trustworthy LLMs through Chinese Textual Ambiguity

2025-07-30 · Xinwei Wu, Haojie Li, Hongyu Liu, Xinyu Ji, Ruohan Li, Yule Chen, Yigeng Zhang arxiv

In this work, we study a critical research problem regarding the trustworthiness of large language models (LLMs): how LLMs behave when encountering ambiguous narrative text, with a particular focus on Chinese textual ambiguity. We created a benchmark dataset by collecting and generating ambiguous sentences with context and their corresponding disambiguated pairs, representing multiple possible interpretations. These annotated examples are systematically categorized into 3 main categories and 9 subcategories. Through experiments, we discovered significant fragility in LLMs when handling ambiguity, revealing behavior that differs substantially from humans. Specifically, LLMs cannot reliably distinguish ambiguous text from unambiguous text, show overconfidence in interpreting ambiguous text as having a single meaning rather than multiple meanings, and exhibit overthinking when attempting to understand the various possible meanings. Our findings highlight a fundamental limitation in current LLMs that has significant implications for their deployment in real-world applications where linguistic ambiguity is common, calling for improved approaches to handle uncertainty in language understanding. The dataset and code are publicly available at this GitHub repository: https://github.com/ictup/LLM-Chinese-Textual-Disambiguation.

📄 PDF Abstract BibTeX arXiv:2507.23121

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Antifragility for Intelligent Autonomous Systems

2018-02-26 · Anusha Mujumdar, Swarup Kumar Mohalik, Ramamurthy Badrinath

Antifragile systems grow measurably better in the presence of hazards. This is in contrast to fragile systems which break down in the presence of hazards, robust systems that tolerate hazards up to a certain degree, and …

Knowledge Mechanisms in Large Language Models: A Survey and Perspective

2024-07-22 · Mengru Wang, Yunzhi Yao, Ziwen Xu, Shuofei Qiao 외

Understanding knowledge mechanisms in Large Language Models (LLMs) is crucial for advancing towards trustworthy AGI. This paper reviews knowledge mechanism analysis from a novel taxonomy including knowledge utilization a…

MemorizationSurvey

Following the Whispers of Values: Unraveling Neural Mechanisms Behind Value-Oriented Behaviors in LLMs

2025-04-07 · Ling Hu, Yuemei Xu, Xiaoyang Gu, Letao Han

Despite the impressive performance of large language models (LLMs), they can present unintended biases and harmful behaviors driven by encoded values, emphasizing the urgent need to understand the value mechanisms behind…

Decision Making

Hidden in the Multiplicative Interaction: Uncovering Fragility in Multimodal Contrastive Learning

2026-04-07 · Tillmann Rheude, Stefan Hegselmann, Roland Eils, Benjamin Wild arxiv

Contrastive learning has become a standard approach for unsupervised learning from paired data, as demonstrated by CLIP for image-text matching. However, many domains involve more than two modalities and require objectiv…

Cross-Modal RetrievalContrastive LearningImage-text matching

CHBench: A Chinese Dataset for Evaluating Health in Large Language Models

2024-09-24 · Chenlu Guo, Nuo Xu, Yi Chang, Yuan Wu

With the rapid development of large language models (LLMs), assessing their performance on health-related inquiries has become increasingly essential. It is critical that these models provide accurate and trustworthy hea…

Misinformation