paper-with-me

Papers

RICA: Evaluating Robust Inference Capabilities Based on Commonsense Axioms

2020-05-02 · EMNLP 2021 11 · Pei Zhou, Rahul Khanna, Seyeon Lee, Bill Yuchen Lin, Daniel Ho, Jay Pujara, Xiang Ren

Pre-trained language models (PTLMs) have achieved impressive performance on commonsense inference benchmarks, but their ability to employ commonsense to make robust inferences, which is crucial for effective communications with humans, is debated. In the pursuit of advancing fluid human-AI communication, we propose a new challenge, RICA: Robust Inference capability based on Commonsense Axioms, that evaluates robust commonsense inference despite textual perturbations. To generate data for this challenge, we develop a systematic and scalable procedure using commonsense knowledge bases and probe PTLMs across two different evaluation settings. Extensive experiments on our generated probe sets with more than 10k statements show that PTLMs perform no better than random guessing on the zero-shot setting, are heavily impacted by statistical biases, and are not robust to perturbation attacks. We also find that fine-tuning on similar statements offer limited gains, as PTLMs still fail to generalize to unseen inferences. Our new large-scale benchmark exposes a significant gap between PTLMs and human-level language understanding and offers a new challenge for PTLMs to demonstrate commonsense.

📄 PDF Abstract BibTeX arXiv:2005.00782

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Filling the Gap: Is Commonsense Knowledge Generation useful for Natural Language Inference?

2025-07-20 · Chathuri Jayaweera, Brianna Yanqui, Bonnie Dorr arxiv

Natural Language Inference (NLI) is the task of determining whether a premise entails, contradicts, or is neutral with respect to a given hypothesis. The task is often framed as emulating human inferential processes, in …

Natural Language Inference

Story Generation with Commonsense Knowledge Graphs and Axioms

2021-09-03 · AKBC Workshop CSKB 2021 10 · Filip Ilievski, Jay Pujara, Hanzhi Zhang

Humans can understand stories, and the rich interactions between agents, locations, and events, seamlessly. However, state-of-the-art reasoning models struggle with understanding, completing, or explaining stories, often…

Common Sense ReasoningKnowledge GraphsStory Generation

Controlling Search in Very large Commonsense Knowledge Bases: A Machine Learning Approach

2016-03-14 · Abhishek Sharma, Michael Witbrock, Keith Goolsbey

Very large commonsense knowledge bases (KBs) often have thousands to millions of axioms, of which relatively few are relevant for answering any given query. A large number of irrelevant axioms can easily overwhelm resolu…

BIG-bench Machine Learning

CIKQA: Learning Commonsense Inference with a Unified Knowledge-in-the-loop QA Paradigm

2022-10-12 · Hongming Zhang, Yintong Huo, Yanai Elazar, Yangqiu Song 외

Recently, the community has achieved substantial progress on many commonsense reasoning benchmarks. However, it is still unclear what is learned from the training process: the knowledge, inference capability, or both? We…

Question AnsweringTask 2

COREALMLIB: An ALM Library Translated from the Component Library

2016-08-06 · Daniela Inclezan

This paper presents COREALMLIB, an ALM library of commonsense knowledge about dynamic domains. The library was obtained by translating part of the COMPONENT LIBRARY (CLIB) into the modular action language ALM. CLIB consi…

Natural Language UnderstandingTranslation