paper-with-me

홈 › Papers

Vera: A General-Purpose Plausibility Estimation Model for Commonsense Statements

2023-05-05 · Jiacheng Liu, Wenya Wang, Dianzhuo Wang, Noah A. Smith, Yejin Choi, Hannaneh Hajishirzi

Despite the much discussed capabilities of today's language models, they are still prone to silly and unexpected commonsense failures. We consider a retrospective verification approach that reflects on the correctness of LM outputs, and introduce Vera, a general-purpose model that estimates the plausibility of declarative statements based on commonsense knowledge. Trained on ~7M commonsense statements created from 19 QA datasets and two large-scale knowledge bases, and with a combination of three training objectives, Vera is a versatile model that effectively separates correct from incorrect statements across diverse commonsense domains. When applied to solving commonsense problems in the verification format, Vera substantially outperforms existing models that can be repurposed for commonsense verification, and it further exhibits generalization capabilities to unseen tasks and provides well-calibrated outputs. We find that Vera excels at filtering LM-generated commonsense knowledge and is useful in detecting erroneous commonsense statements generated by models like ChatGPT in real-world settings.

📄 PDF Abstract BibTeX arXiv:2305.03695

Code (1)

liujch1998/vera 공식 구현 pytorch

Similar Papers 제목 키워드 기반

Estimating Commonsense Plausibility through Semantic Shifts

2025-02-19 · Wanqing Cui, Keping Bi, Jiafeng Guo, Xueqi Cheng

Commonsense plausibility estimation is critical for evaluating language models (LMs), yet existing generative approaches--reliant on likelihoods or verbalized judgments--struggle with fine-grained discrimination. In this…

KEPR: Knowledge Enhancement and Plausibility Ranking for Generative Commonsense Question Answering

2023-05-15 · Zhifeng Li, Bowei Zou, Yifan Fan, Yu Hong

Generative commonsense question answering (GenCQA) is a task of automatically generating a list of answers given a question. The answer list is required to cover all reasonable answers. This presents the considerable cha…

Passage RetrievalQuestion AnsweringRetrieval

Holistic++ Scene Understanding: Single-view 3D Holistic Scene Parsing and Human Pose Estimation with Human-Object Interaction and Physical Commonsense

2019-09-04 · ICCV 2019 10 · Yixin Chen, Siyuan Huang, Tao Yuan, Siyuan Qi 외

We propose a new 3D holistic++ scene understanding problem, which jointly tackles two tasks from a single-view image: (i) holistic scene parsing and reconstruction---3D estimations of object bounding boxes, camera pose, …

3D Human Pose EstimationHuman-Object Interaction DetectionPose EstimationScene Parsing+1

SituatedGen: Incorporating Geographical and Temporal Contexts into Generative Commonsense Reasoning

2023-06-21 · NeurIPS 2023 11 · Yunxiang Zhang, Xiaojun Wan

Recently, commonsense reasoning in text generation has attracted much attention. Generative commonsense reasoning is the task that requires machines, given a group of keywords, to compose a single coherent sentence with …

SentenceText Generation

EntailE: Introducing Textual Entailment in Commonsense Knowledge Graph Completion

2024-02-15 · Ying Su, Tianqing Fang, Huiru Xiao, Weiqi Wang 외

Commonsense knowledge graph completion is a new challenge for commonsense knowledge graph construction and application. In contrast to factual knowledge graphs such as Freebase and YAGO, commonsense knowledge graphs (CSK…

graph constructionGraph EmbeddingKnowledge Graph CompletionKnowledge Graph Embedding+2