paper-with-me

홈 › Papers

Do language models have coherent mental models of everyday things?

2022-12-20 · Yuling Gu, Bhavana Dalvi Mishra, Peter Clark

When people think of everyday things like an egg, they typically have a mental image associated with it. This allows them to correctly judge, for example, that "the yolk surrounds the shell" is a false statement. Do language models similarly have a coherent picture of such everyday things? To investigate this, we propose a benchmark dataset consisting of 100 everyday things, their parts, and the relationships between these parts, expressed as 11,720 "X relation Y?" true/false questions. Using these questions as probes, we observe that state-of-the-art pre-trained language models (LMs) like GPT-3 and Macaw have fragments of knowledge about these everyday things, but do not have fully coherent "parts mental models" (54-59% accurate, 19-43% conditional constraint violation). We propose an extension where we add a constraint satisfaction layer on top of the LM's raw predictions to apply commonsense constraints. As well as removing inconsistencies, we find that this also significantly improves accuracy (by 16-20%), suggesting how the incoherence of the LM's pictures of everyday things can be significantly reduced.

📄 PDF Abstract BibTeX arXiv:2212.10029

Code (1)

allenai/everyday-things 공식 구현

Methods 이 논문이 사용한 방법론

{Dispute@FaQ-s}How to file a dispute with Expedia? How to file a dispute with Expedia? To file a complaint against Expedia, first try contacting their customer service directly. You can reach them by phone at…
Multi-Head Attention 설명 없음
Attention 설명 없음
Inverse Square Root Schedule Inverse Square Root is a learning rate schedule 1 / $\sqrt{\max\left(n, k\right)}$ where $n$ is the current training iteration and $k$ is the number of warm-up steps. This…
Weight Decay 설명 없음
15 Ways to Contact How can i speak to someone at Delta Airlines 설명 없음
Cosine Annealing Cosine Annealing is a type of learning rate schedule that has the effect of starting with a large learning rate that is relatively rapidly decreased to a minimum value before…
Adafactor Adafactor is a stochastic optimization method based on Adam that reduces memory usage while retaining the empirical benefits of…

Similar Papers 제목 키워드 기반

The logic behind desirable sets of things, and its filter representation

2023-02-16 · Gert de Cooman, Arthur Van Camp, Jasper De Bock

We identify the (filter representation of the) logic behind the recent theory of coherent sets of desirable (sets of) things, which generalise coherent sets of desirable (sets of) gambles as well as coherent choice funct…

Thoughtful Things: Building Human-Centric Smart Devices with Small Language Models

2024-05-06 · Evan King, Haoxiang Yu, Sahil Vartak, Jenna Jacob 외

Everyday devices like light bulbs and kitchen appliances are now embedded with so many features and automated behaviors that they have become complicated to actually use. While such "smart" capabilities can better suppor…

Internet of Things Device Capabilities, Architectures, Protocols, and Smart Applications in Healthcare Domain: A Review

2022-04-12 · Md. Milon Islam, Sheikh Nooruddin, Fakhri Karray, Ghulam Muhammad

Nowadays, the Internet has spread to practically every country around the world and is having unprecedented effects on people's lives. The Internet of Things (IoT) is getting more popular and has a high level of interest…

Semantic Reasoning for Context-aware Internet of Things Applications

2016-04-28 · Altti Ilari Maarala, Xiang Su, Jukka Riekki

Advances in ICT are bringing into reality the vision of a large number of uniquely identifiable, interconnected objects and things that gather information from diverse physical environments and deliver the information to…

Context-Aware Stress Monitoring using Wearable and Mobile Technologies in Everyday Settings

2023-12-14 · Seyed Amir Hossein Aqajari, Sina Labbaf, Phuc Hoang Tran, Brenda Nguyen 외

Daily monitoring of stress is a critical component of maintaining optimal physical and mental health. Physiological signals and contextual information have recently emerged as promising indicators for detecting instances…