paper-with-me

홈 › Papers

Beyond Transformers for Function Learning

2023-04-19 · Simon Segert, Jonathan Cohen

The ability to learn and predict simple functions is a key aspect of human intelligence. Recent works have started to explore this ability using transformer architectures, however it remains unclear whether this is sufficient to recapitulate the extrapolation abilities of people in this domain. Here, we propose to address this gap by augmenting the transformer architecture with two simple inductive learning biases, that are directly adapted from recent models of abstract reasoning in cognitive science. The results we report demonstrate that these biases are helpful in the context of large neural network models, as well as shed light on the types of inductive learning biases that may contribute to human abilities in extrapolation.

📄 PDF Abstract BibTeX arXiv:2304.09979

Code (0)

등록된 구현이 없습니다.

Tasks

Inductive Learning

Similar Papers 제목 키워드 기반

Transformers Meet In-Context Learning: A Universal Approximation Theory

2025-06-05 · Gen Li, Yuchen Jiao, Yu Huang, Yuting Wei 외

Modern large language models are capable of in-context learning, the ability to perform new tasks at inference time using only a handful of input-output examples in the prompt, without any fine-tuning or parameter update…

In-Context Learning

Optimality and NP-Hardness of Transformers in Learning Markovian Dynamical Functions

2025-10-21 · Yanna Ding, Songtao Lu, Yingdong Lu, Tomasz Nowicki 외 arxiv

Transformer architectures can solve unseen tasks based on input-output pairs in a given prompt due to in-context learning (ICL). Existing theoretical studies on ICL have mainly focused on linear regression tasks, often w…

How Do Transformers Learn In-Context Beyond Simple Functions? A Case Study on Learning with Representations

2023-10-16 · Tianyu Guo, Wei Hu, Song Mei, Huan Wang 외

While large language models based on the transformer architecture have demonstrated remarkable in-context learning (ICL) capabilities, understandings of such capabilities are still in an early stage, where existing theor…

In-Context Learning

Representation Matters for Mastering Chess: Improved Feature Representation in AlphaZero Outperforms Switching to Transformers

2023-04-28 · Johannes Czech, Jannis Blüml, Kristian Kersting, Hedinn Steingrimsson

While transformers have gained recognition as a versatile tool for artificial intelligence (AI), an unexplored challenge arises in the context of chess - a classical AI benchmark. Here, incorporating Vision Transformers …

Game of Chess

Cooperation Is All You Need

2023-05-16 · Ahsan Adeel, Junaid Muzaffar, Fahad Zia, Khubaib Ahmed 외

Going beyond 'dendritic democracy', we introduce a 'democracy of local processors', termed Cooperator. Here we compare their capabilities when used in permutation invariant neural networks for reinforcement learning (RL)…

AllReinforcement Learning (RL)