paper-with-me

홈 › Papers

A Comprehensive Evaluation of Semantic Relation Knowledge of Pretrained Language Models and Humans

2024-12-02 · Zhihan Cao, Hiroaki Yamada, Simone Teufel, Takenobu Tokunaga

Recently, much work has concerned itself with the enigma of what exactly PLMs (pretrained language models) learn about different aspects of language, and how they learn it. One stream of this type of research investigates the knowledge that PLMs have about semantic relations. However, many aspects of semantic relations were left unexplored. Only one relation was considered, namely hypernymy. Furthermore, previous work did not measure humans' performance on the same task as that solved by the PLMs. This means that at this point in time, there is only an incomplete view of models' semantic relation knowledge. To address this gap, we introduce a comprehensive evaluation framework covering five relations beyond hypernymy, namely hyponymy, holonymy, meronymy, antonymy, and synonymy. We use six metrics (two newly introduced here) for recently untreated aspects of semantic relation knowledge, namely soundness, completeness, symmetry, asymmetry, prototypicality, and distinguishability and fairly compare humans and models on the same task. Our extensive experiments involve 16 PLMs, eight masked and eight causal language models. Up to now only masked language models had been tested although causal and masked language models treat context differently. Our results reveal a significant knowledge gap between humans and models for almost all semantic relations. Antonymy is the outlier relation where all models perform reasonably well. In general, masked language models perform significantly better than causal language models. Nonetheless, both masked and causal language models are likely to confuse non-antonymy relations with antonymy.

📄 PDF Abstract BibTeX arXiv:2412.01131

Code (0)

등록된 구현이 없습니다.

Tasks

Relation

Similar Papers 제목 키워드 기반

Probing Pretrained Language Models with Hierarchy Properties

2023-12-15 · Jesús Lovón-Melgarejo, Jose G. Moreno, Romaric Besançon, Olivier Ferret 외

Since Pretrained Language Models (PLMs) are the cornerstone of the most recent Information Retrieval (IR) models, the way they encode semantic knowledge is particularly important. However, little attention has been given…

Hypernym DiscoveryInformation RetrievalReading ComprehensionRetrieval

InfoCLIP: Bridging Vision-Language Pretraining and Open-Vocabulary Semantic Segmentation via Information-Theoretic Alignment Transfer

2025-11-20 · Muyao Yuan, Yuanhong Zhang, Weizhan Zhang, Lan Ma 외 arxiv

Recently, the strong generalization ability of CLIP has facilitated open-vocabulary semantic segmentation, which labels pixels using arbitrary text. However, existing methods that fine-tune CLIP for segmentation on limit…

Semantic Segmentation

Delving Deep into Semantic Relation Distillation

2025-03-27 · Zhaoyi Yan, KangJun Liu, Qixiang Ye

Knowledge distillation has become a cornerstone technique in deep learning, facilitating the transfer of knowledge from complex models to lightweight counterparts. Traditional distillation approaches focus on transferrin…

Knowledge DistillationModel CompressionRelationSuperpixels

Relative Counterfactual Contrastive Learning for Mitigating Pretrained Stance Bias in Stance Detection

2024-05-16 · Jiarui Zhang, Shaojuan Wu, Xiaowang Zhang, Zhiyong Feng

Stance detection classifies stance relations (namely, Favor, Against, or Neither) between comments and targets. Pretrained language models (PLMs) are widely used to mine the stance relation to improve the performance of …

Contrastive LearningcounterfactualLanguage ModelingLanguage Modelling+2

Can Linguistic Knowledge Improve Multimodal Alignment in Vision-Language Pretraining?

2023-08-24 · Fei Wang, Liang Ding, Jun Rao, Ye Liu 외

The multimedia community has shown a significant interest in perceiving and representing the physical world with multimodal pretrained neural network models, and among them, the visual-language pertaining (VLP) is, curre…

AttributeNegationSentence