paper-with-me

홈 › Papers

Benchmarks for Automated Commonsense Reasoning: A Survey

2023-02-09 · Ernest Davis

More than one hundred benchmarks have been developed to test the commonsense knowledge and commonsense reasoning abilities of artificial intelligence (AI) systems. However, these benchmarks are often flawed and many aspects of common sense remain untested. Consequently, we do not currently have any reliable way of measuring to what extent existing AI systems have achieved these abilities. This paper surveys the development and uses of AI commonsense benchmarks. We discuss the nature of common sense; the role of common sense in AI; the goals served by constructing commonsense benchmarks; and desirable features of commonsense benchmarks. We analyze the common flaws in benchmarks, and we argue that it is worthwhile to invest the work needed ensure that benchmark examples are consistently high quality. We survey the various methods of constructing commonsense benchmarks. We enumerate 139 commonsense benchmarks that have been developed: 102 text-based, 18 image-based, 12 video based, and 7 simulated physical environments. We discuss the gaps in the existing benchmarks and aspects of commonsense reasoning that are not addressed in any existing benchmark. We conclude with a number of recommendations for future development of commonsense AI benchmarks.

📄 PDF Abstract BibTeX arXiv:2302.04752

Code (0)

등록된 구현이 없습니다.

Tasks

Common Sense ReasoningSurvey

Methods 이 논문이 사용한 방법론

Test 설명 없음

Similar Papers 제목 키워드 기반

Natural Language Processing with Commonsense Knowledge: A Survey

2021-08-10 · Yubo Xie, Zonghui Liu, Zongyang Ma, Fanyuan Meng 외

Commonsense knowledge is essential for advancing natural language processing (NLP) by enabling models to engage in human-like reasoning, which requires a deeper understanding of context and often involves making inferenc…

Survey

Commonsense Reasoning for Conversational AI: A Survey of the State of the Art

2023-02-15 · Christopher Richardson, Larry Heck

Large, transformer-based pretrained language models like BERT, GPT, and T5 have demonstrated a deep understanding of contextual semantics and language syntax. Their success has enabled significant advances in conversatio…

What Really is Commonsense Knowledge?

2024-11-06 · Quyet V. Do, Junze Li, Tung-Duong Vuong, Zhaowei Wang 외

Commonsense datasets have been well developed in Natural Language Processing, mainly through crowdsource human annotation. However, there are debates on the genuineness of commonsense reasoning benchmarks. In specific, a…

The Odyssey of Commonsense Causality: From Foundational Benchmarks to Cutting-Edge Reasoning

2024-06-27 · Shaobo Cui, Zhijing Jin, Bernhard Schölkopf, Boi Faltings

Understanding commonsense causality is a unique mark of intelligence for humans. It helps people understand the principles of the real world better and benefits the decision-making process related to causation. For insta…

ArticlesDecision Making

Commonsense Knowledge Reasoning and Generation with Pre-trained Language Models: A Survey

2022-01-28 · Prajjwal Bhargava, Vincent Ng

While commonsense knowledge acquisition and reasoning has traditionally been a core research topic in the knowledge representation and reasoning community, recent years have seen a surge of interest in the natural langua…