paper-with-me

홈 › Papers

Modeling Open-World Cognition as On-Demand Synthesis of Probabilistic Models

2025-07-16 · Lionel Wong, Katherine M. Collins, Lance Ying, Cedegao E. Zhang, Adrian Weller, Tobias Gerstenberg, Timothy O'Donnell, Alexander K. Lew, Jacob D. Andreas, Joshua B. Tenenbaum, Tyler Brooke-Wilson arxiv

When faced with novel situations, people are able to marshal relevant considerations from a wide range of background knowledge and put these to use in inferences and predictions. What permits us to draw in globally relevant information and reason over it coherently? Here, we explore the hypothesis that people use a combination of distributed and symbolic representations to construct bespoke mental models tailored to novel situations. We propose a computational implementation of this idea -- a `Model Synthesis Architecture'' (MSA) -- using language models to implement global relevance-based retrieval and model synthesis and probabilistic programs to implement bespoke, coherent world models. We evaluate our MSA as a model of human judgments on a novel reasoning dataset. The dataset -- built around a Model Olympics` domain of sports vignettes -- tests models' capacity for human-like, open-ended reasoning by requiring (i) judgments about novel causal structures described in language; (ii) drawing on large bodies of background knowledge; and (iii) doing both in light of observations that introduce arbitrary novel variables. Our MSA approach captures human judgments better than language model-only baselines, under both direct and chain-of-thought generations from the LM that supports model synthesis. These results suggest that MSAs can be implemented in a way that mirrors people's ability to deliver locally coherent reasoning over globally relevant variables, offering a path to understanding and replicating human reasoning in open-ended domains.

📄 PDF Abstract BibTeX arXiv:2507.12547

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

ENACT: Evaluating Embodied Cognition with World Modeling of Egocentric Interaction

2025-11-26 · Qineng Wang, Wenlong Huang, Yu Zhou, Hang Yin 외 arxiv

Embodied cognition argues that intelligence arises from sensorimotor interaction rather than passive observation. It raises an intriguing question: do modern vision-language models (VLMs), trained largely in a disembodie…

Visual Question AnsweringAffordance Recognition

TTS-1 Technical Report

2025-07-22 · Oleg Atamanenko, Anna Chalova, Joseph Coombes, Nikki Cope 외 arxiv

We introduce Inworld TTS-1, a set of two Transformer-based autoregressive text-to-speech (TTS) models. Our largest model, TTS-1-Max, has 8.8B parameters and is designed for utmost quality and expressiveness in demanding …

Speech Synthesis

MuSHRoom: Multi-Sensor Hybrid Room Dataset for Joint 3D Reconstruction and Novel View Synthesis

2023-11-05 · Xuqian Ren, Wenjia Wang, Dingding Cai, Tuuli Tuominen 외

Metaverse technologies demand accurate, real-time, and immersive modeling on consumer-grade hardware for both non-human perception (e.g., drone/robot/autonomous car navigation) and immersive technologies like AR/VR, requ…

3D ReconstructionNovel View Synthesis

Exploring the Open World Using Incremental Extreme Value Machines

2022-05-30 · Tobias Koch, Felix Liebezeit, Christian Riess, Vincent Christlein 외

Dynamic environments require adaptive applications. One particular machine learning problem in dynamic environments is open world recognition. It characterizes a continuously changing domain where only some classes are s…

Class Incremental LearningComputational EfficiencyFace Recognitionimage-classification+3

PoE-World: Compositional World Modeling with Products of Programmatic Experts

2025-05-16 · Wasu Top Piriyakulkij, Yichao Liang, Hao Tang, Adrian Weller 외

Learning how the world works is central to building AI agents that can adapt to complex environments. Traditional world models based on deep learning demand vast amounts of training data, and do not flexibly update their…

Montezuma's RevengeProgram Synthesis