paper-with-me

홈 › Papers

When do neural networks learn world models?

2025-02-13 · Tianren Zhang, GuanYu Chen, Feng Chen

Humans develop world models that capture the underlying generation process of data. Whether neural networks can learn similar world models remains an open problem. In this work, we provide the first theoretical results for this problem, showing that in a multi-task setting, models with a low-degree bias provably recover latent data-generating variables under mild assumptions -- even if proxy tasks involve complex, non-linear functions of the latents. However, such recovery is also sensitive to model architecture. Our analysis leverages Boolean models of task solutions via the Fourier-Walsh transform and introduces new techniques for analyzing invertible Boolean transforms, which may be of independent interest. We illustrate the algorithmic implications of our results and connect them to related research areas, including self-supervised learning, out-of-distribution generalization, and the linear representation hypothesis in large language models.

📄 PDF Abstract BibTeX arXiv:2502.09297

Code (0)

등록된 구현이 없습니다.

Tasks

Out-of-Distribution GeneralizationSelf-Supervised Learning

Similar Papers 제목 키워드 기반

Novelty Detection in Reinforcement Learning with World Models

2023-10-12 · Geigh Zollicoffer, Kenneth Eaton, Jonathan Balloch, Julia Kim 외

Reinforcement learning (RL) using world models has found significant recent successes. However, when a sudden change to world mechanics or properties occurs then agent performance and reliability can dramatically decline…

Decision MakingNovelty Detectionreinforcement-learningReinforcement Learning+1

Sample Efficient Robot Learning with Structured World Models

2022-10-21 · Tuluhan Akbulut, Max Merlin, Shane Parr, Benedict Quartey 외

Reinforcement learning has been demonstrated as a flexible and effective approach for learning a range of continuous control tasks, such as those used by robots to manipulate objects in their environment. But in robotics…

continuous-controlContinuous Control

A Retrospective on the Robot Air Hockey Challenge: Benchmarking Robust, Reliable, and Safe Learning Techniques for Real-world Robotics

2024-11-08 · Puze Liu, Jonas Günster, Niklas Funk, Simon Gröger 외

Machine learning methods have a groundbreaking impact in many application domains, but their application on real robotic platforms is still limited. Despite the many challenges associated with combining machine learning …

Benchmarking

When and How Much to Imagine: Adaptive Test-Time Scaling with World Models for Visual Spatial Reasoning

2026-02-09 · Shoubin Yu, Yue Zhang, Zun Wang, Jaehong Yoon 외 arxiv

Despite rapid progress in MLLMs, visual spatial reasoning remains unreliable when correct answers depend on how a scene would appear under unseen or alternative viewpoints. Recent work addresses this by augmenting reason…

Spatial Reasoning

Changing Simplistic Worldviews

2024-01-05 · Maxim Senkov, Toygar T. Kerman

We study a Bayesian persuasion model with two-dimensional states of the world, in which the sender (she) and receiver (he) have heterogeneous prior beliefs and care about different dimensions. The receiver is a naive age…