paper-with-me

Papers

Cross-Document Cross-Lingual NLI via RST-Enhanced Graph Fusion and Interpretability Prediction

2025-04-11 · Mengying Yuan, Wenhao Wang, Zixuan Wang, Yujie Huang, Kangli Wei, Fei Li, Chong Teng, Donghong Ji

Natural Language Inference (NLI) is a fundamental task in natural language processing. While NLI has developed many sub-directions such as sentence-level NLI, document-level NLI and cross-lingual NLI, Cross-Document Cross-Lingual NLI (CDCL-NLI) remains largely unexplored. In this paper, we propose a novel paradigm: CDCL-NLI, which extends traditional NLI capabilities to multi-document, multilingual scenarios. To support this task, we construct a high-quality CDCL-NLI dataset including 25,410 instances and spanning 26 languages. To address the limitations of previous methods on CDCL-NLI task, we further propose an innovative method that integrates RST-enhanced graph fusion with interpretability-aware prediction. Our approach leverages RST (Rhetorical Structure Theory) within heterogeneous graph neural networks for cross-document context modeling, and employs a structure-aware semantic alignment based on lexical chains for cross-lingual understanding. For NLI interpretability, we develop an EDU (Elementary Discourse Unit)-level attribution framework that produces extractive explanations. Extensive experiments demonstrate our approach's superior performance, achieving significant improvements over both conventional NLI models as well as large language models. Our work sheds light on the study of NLI and will bring research interest on cross-document cross-lingual context understanding, hallucination elimination and interpretability inference. Our code and datasets are available at \href{https://anonymous.4open.science/r/CDCL-NLI-637E/}{CDCL-NLI-link} for peer review.

📄 PDF Abstract BibTeX arXiv:2504.12324

Code (0)

등록된 구현이 없습니다.

Tasks

Cross-Lingual Natural Language InferenceGraph AttentionHallucinationInformation RetrievalNatural Language InferenceRetrievalSemantic Retrieval

Similar Papers 제목 키워드 기반

Zero-Shot Cross-Lingual Document-Level Event Causality Identification with Heterogeneous Graph Contrastive Transfer Learning

2024-03-05 · Zhitao He, Pengfei Cao, Zhuoran Jin, Yubo Chen 외

Event Causality Identification (ECI) refers to the detection of causal relations between events in texts. However, most existing studies focus on sentence-level ECI with high-resource languages, leaving more challenging …

Event Causality IdentificationFew-Shot LearningSentenceTransfer Learning

Cross-lingual Text Classification with Heterogeneous Graph Neural Network

2021-05-24 · ACL 2021 5 · ZiYun Wang, Xuan Liu, Peiji Yang, Shixing Liu 외

Cross-lingual text classification aims at training a classifier on the source language and transferring the knowledge to target languages, which is very useful for low-resource languages. Recent multilingual pretrained l…

ClassificationGraph Neural NetworkSemantic SimilaritySemantic Textual Similarity+2

A Contextual Alignment Enhanced Cross Graph Attention Network for Cross-lingual Entity Alignment

2020-12-01 · COLING 2020 8 · Zhiwen Xie, Runjie Zhu, Kunsong Zhao, Jin Liu 외

Cross-lingual entity alignment, which aims to match equivalent entities in KGs with different languages, has attracted considerable focus in recent years. Recently, many graph neural network (GNN) based methods are propo…

Entity AlignmentGraph AttentionGraph Neural Network

Cross-lingual Data Augmentation for Document-grounded Dialog Systems in Low Resource Languages

2023-05-24 · Qi Gou, Zehua Xia, Wenzhe Du

This paper proposes a framework to address the issue of data scarcity in Document-Grounded Dialogue Systems(DGDS). Our model leverages high-resource languages to enhance the capability of dialogue generation in low-resou…

Data AugmentationDecoderDialogue GenerationRetrieval

LuxEmbedder: A Cross-Lingual Approach to Enhanced Luxembourgish Sentence Embeddings

2024-12-04 · Fred Philippy, Siwen Guo, Jacques Klein, Tegawendé F. Bissyandé

Sentence embedding models play a key role in various Natural Language Processing tasks, such as in Topic Modeling, Document Clustering and Recommendation Systems. However, these models rely heavily on parallel data, whic…

Recommendation SystemsSentenceSentence EmbeddingSentence-Embedding+1