paper-with-me

홈 › Papers

Learning to Chain Operations by Routing Information Through a Global Workspace

2025-02-28 · Hugo Chateau-Laurent, Rufin VanRullen

We present a model inspired by the Global Workspace Theory that integrates specialized modules to perform a sequential reasoning task. A controller selectively routes information between modules through the workspace using a gating mechanism. This approach allows the model to chain operations by iteratively broadcasting information between specialized domains, mimicking System-2 reasoning. We evaluate the model's performance on a simple addition task, where two addends must be summed. The task can be solved by routing information sequentially through an Input module, an Increment module (multiple times), and finally an Output module. We consider two implementations of this system with increasing complexity. First, using hand-designed modules operating on one-hot digit representations, the controller (a LSTM recurrent network) learns to select the appropriate modules (input, increment, output) in the appropriate sequence. Second, we replace the hand-designed modules with learned representation modules for MNIST images and an increment module trained on the task objectives; here again, the controller learns the appropriate sequential module selection to solve the task. Finally, we show that the Global Workspace model, while having fewer parameters, outperforms LSTMs and Transformers when tested on unseen addition operations (both interpolations and extrapolations of addition operations seen during training). Our results highlight the potential of architectures inspired by the Global Workspace Theory to enhance deep learning's reasoning capabilities.

📄 PDF Abstract BibTeX arXiv:2503.01906

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Sigmoid Activation 설명 없음
Tanh Activation 설명 없음
LSTM An LSTM is a type of recurrent neural network that addresses the vanishing gradient problem in vanilla…

Similar Papers 제목 키워드 기반

Using Reinforcement Learning for the Three-Dimensional Loading Capacitated Vehicle Routing Problem

2023-07-22 · Stefan Schoepf, Stephen Mak, Julian Senoner, Liming Xu 외

Heavy goods vehicles are vital backbones of the supply chain delivery system but also contribute significantly to carbon emissions with only 60% loading efficiency in the United Kingdom. Collaborative vehicle routing has…

reinforcement-learningReinforcement Learning

MACRO: Markov Chain Routing of Transformer Layers

2026-08-06 · Paweł Batorski, Abtin Pourhadi, Akylgali Aitaza, Przemysław Spurek 외 arxiv

Standard Large Language Models (LLMs) execute layers sequentially. Dynamic layer routing, i.e. search for a different execution path through layers involving layer repetitions, skips and other moves, can improve performa…

SCOPE: Supply-Chain Operations through Coupled Policies for End-to-End Coordination

2026-07-30 · Yunhao Liang, Xianqi Cao, Pujun Zhang, Yuan Qu 외 arxiv

Can supply-chain AI move beyond isolated decision modules toward unified operational planning? A complete replenishment plan specifies which products each location carries, which upstream facility supplies it, how often …

Virtual Quantum Markov Chains

2023-12-04 · Yu-Ao Chen, Chengkai Zhu, Keming He, Mingrui Jing 외

Quantum Markov chains generalize classical Markov chains for random variables to the quantum realm and exhibit unique inherent properties, making them an important feature in quantum information theory. In this work, we …

Reinforcement Learning for Multi-Truck Vehicle Routing Problems

2022-11-30 · Joshua Levin, Randall Correll, Takanori Ide, Takafumi Suzuki 외

Deep reinforcement learning (RL) has been shown to be effective in producing approximate solutions to some vehicle routing problems (VRPs), especially when using policies generated by encoder-decoder attention mechanisms…

Combinatorial OptimizationDecoderDeep Reinforcement Learningreinforcement-learning+2