RuleBert: Teaching Soft Rules to Pre-trained Language Models
While pre-trained language models (PLMs) are the go-to solution to tackle many natural language processing problems, they are still very limited in their ability to capture and to use common-sense knowledge. In fact, even if information is available in the form of approximate (soft) logical rules, it is not clear how to transfer it to a PLM in order to improve its performance for deductive reasoning tasks. Here, we aim to bridge this gap by teaching PLMs how to reason with soft Horn rules. We introduce a classification task where, given facts and soft rules, the PLM should return a prediction with a probability for a given hypothesis. We release the first dataset for this task, and we propose a revised loss function that enables the PLM to learn how to predict precise probabilities for the task. Our evaluation results show that the resulting fine-tuned models achieve very high performance, even on logical rules that were unseen at training. Moreover, we demonstrate that logical notions expressed by the rules are transferred to the fine-tuned model, yielding state-of-the-art results on external datasets.
Code (1)
Tasks
Common Sense ReasoningSimilar Papers 제목 키워드 기반
Distilling Task-specific Logical Rules from Large Pre-trained Models
Logical rules, both transferable and explainable, are widely used as weakly supervised signals for many downstream tasks such as named entity tagging. To reduce the human effort of writing rules, previous researchers ado…
Large Language Model-Driven Classroom Flipping: Empowering Student-Centric Peer Questioning with Flipped Interaction
Reciprocal questioning is essential for effective teaching and learning, fostering active engagement and deeper understanding through collaborative interactions, especially in large classrooms. Can large language model (…
ChatbotLanguage ModelingLanguage ModellingLarge Language Model+1Teaching Probabilistic Logical Reasoning to Transformers
In this paper, we evaluate the capability of transformer-based language models in making inferences over uncertain text that includes uncertain rules of reasoning. We cover both Pre-trained Language Models (PLMs) and gen…
Logical ReasoningQuestion AnsweringPāṇinian Phonological Changes: Computation and Development of Online Access System
Pāṇini used the term saṃhitā for phonological changes. Any Sound change which alters phonemes in a particular language is called Phonological Change. It arises when two sounds are pronounced in a language with uninterrup…
SentenceNatural Language Generation for Polysynthetic Languages: Language Teaching and Learning Software for Kanyen'k\'eha (Mohawk)
Kanyen{'}k{\'e}ha (in English, Mohawk) is an Iroquoian language spoken primarily in Eastern Canada (Ontario, Qu{\'e}bec). Classified as endangered, it has only a small number of speakers and very few younger native speak…
Text Generation