paper-with-me

홈 › Papers

Convergent World Representations and Divergent Tasks

2026-01-31 · Core Francisco Park arxiv

While neural representations are central to modern deep learning, the conditions governing their geometry and their roles in downstream adaptability remain poorly understood. We develop a framework clearly separating the underlying world, the data generation process and the resulting model representations to study these questions in a controlled setup. 5,075 city coordinates define the world and 7 geometric tasks generate the training data for autoregressive training. We find that different tasks give rise to qualitatively and quantitatively distinct world representation geometries. However, multi-task training drives convergence of world representations: models trained on non-overlapping tasks develop aligned geometric representations, providing controlled evidence for the Multitask Scaling Hypothesis of the Platonic Representation Hypothesis. To study adaptation, we pretrain models on all tasks, then test whether new entities (cities) can be consistently integrated into the representation space via fine-tuning. Surprisingly, we find that despite multi-task pretraining, some tasks, which we call divergent, actively harm the representational integration of new entities and harm generalization. Our results show that training on multiple relational tasks reliably produces convergent world representations, but lurking divergent tasks can catastrophically harm new entity integration via fine-tuning.

📄 PDF Abstract BibTeX arXiv:2602.00533

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Functional and spatial rewiring jointly generate convergent-divergent units in self-organizing networks

2021-03-30 · Jia Li, Ilias Rentzeperis, Cees van Leeuwen

Self-organization through adaptive rewiring of random neural networks generates brain-like topologies comprising modular small-world structures with rich club effects, merely as the product of optimizing the network topo…

Sensitivity

From Shots to Stories: LLM-Assisted Video Editing with Unified Language Representations

2025-05-18 · Yuzhi Li, Haojun Xu, Fang Tian

Large Language Models (LLMs) and Vision-Language Models (VLMs) have demonstrated remarkable reasoning and generalization capabilities in video understanding; however, their application in video editing remains largely un…

Video EditingVideo Understanding

Reasoning Beyond the Obvious: Evaluating Divergent and Convergent Thinking in LLMs for Financial Scenarios

2025-07-24 · Zhuang Qiang Bok, Watson Wei Khong Chua arxiv

Most reasoning benchmarks for LLMs emphasize factual accuracy or step-by-step logic. In finance, however, professionals must not only converge on optimal decisions but also generate creative, plausible futures under unce…

Adaptive rewiring of random neural networks generates convergent-divergent units

2021-04-03 · Ilias Rentzeperis, Steeve Laquitaine, Cees van Leeuwen

Brain networks are adaptively rewired continually, adjusting their topology to bring about functionality and efficiency in sensory, motor and cognitive tasks. In model neural network architectures, adaptive rewiring gene…

Distributed Computing

The Neural Basis and Evolution of Divergent and Convergent Thought

2019-03-13

This chapter takes as its departure point a neural level theory of insight that arose from studies of the sparse, distributed, content-addressable architecture of associative memory. It is argued that convergent thought …