Can We Do Interpretable NLI with Graphs Based on Atomic Propositions?
While Large Language Model (LLM)-based Natural Language Inference (NLI) systems achieve high accuracy, their decision-making processes lack auditable structures. This paper explores whether NLI can be performed using only interpretable, graph-based representations of evidence. We introduce a fully graph-based pipeline where the classifier never directly processes the input text. Instead, sentences are decomposed into atomic propositions, converted into ConceptNet triples via constrained decoding, and represented as three graphs per pair: premise, hypothesis, and a retrieved ConceptNet subgraph. These graphs are then fed into a fine-tuned 0.8-billion-parameter language model. On the SNLI dataset, our pipeline achieves 89.7% accuracy, just 1.9 points below an identically trained text-based model. On ANLI, it matches the published performance of RoBERTa-large on rounds R2 and R3 (50% accuracy) but trails by 16 points on R1, resulting in an overall gap of 9 to 14 points compared to its text counterpart. We term this gap the price of interpretability and demonstrate that it stems from representational limitations rather than data constraints. Ablation studies further reveal that graphs and text are complementary: combining both modalities achieves 92.1% accuracy on SNLI.
Code (0)
등록된 구현이 없습니다.
Tasks
Natural Language InferenceResults from the Paper
| Rank | Task | Dataset | Model | Metrics |
|---|---|---|---|---|
| #99 | Natural Language Inference | SNLI | Can We Do Interpretable NLI with Graphs | Accuracy: 89.7 |
Similar Papers 제목 키워드 기반
LLM-based Atomic Propositions help weak extractors: Evaluation of a Propositioner for triplet extraction
Knowledge Graph construction from natural language requires extracting structured triplets from complex, information-dense sentences. In this paper, we investigate if the decomposition of text into atomic propositions (m…
Knowledge DistillationRelation ExtractionTask-Oriented Active Perception and Planning in Environments with Partially Known Semantics
We consider an agent that is assigned with a temporal logic task in an environment whose semantic representation is only partially known. We represent the semantics of the environment with a set of state properties, call…
Collaborative rover-copter path planning and exploration with temporal logic specifications based on Bayesian update under uncertain environments
This paper investigates a collaborative rover-copter path planning and exploration with temporal logic specifications under uncertain environments. The objective of the rover is to complete a mission expressed by a synta…
Joint Learning of Reward Machines and Policies in Environments with Partially Known Semantics
We study the problem of reinforcement learning for a task encoded by a reward machine. The task is defined over a set of properties in the environment, called atomic propositions, and represented by Boolean variables. On…
Q-Learningreinforcement-learningReinforcement Learning (RL)PlatoLTL: Learning to Generalize Across Symbols in LTL Instructions for Multi-Task RL
A central challenge in multi-task reinforcement learning (RL) is to train generalist policies capable of performing tasks not seen during training. To facilitate such generalization, linear temporal logic (LTL) has emerg…
Zero-shot GeneralizationReinforcement Learning