paper-with-me

홈 › Papers

The dynamic interplay between in-context and in-weight learning in humans and neural networks

2024-02-13 · Jacob Russin, Ellie Pavlick, Michael J. Frank

Human learning embodies a striking duality: sometimes, we appear capable of following logical, compositional rules and benefit from structured curricula (e.g., in formal education), while other times, we rely on an incremental approach or trial-and-error, learning better from curricula that are randomly interleaved. Influential psychological theories explain this seemingly disparate behavioral evidence by positing two qualitatively different learning systems -- one for rapid, rule-based inferences and another for slow, incremental adaptation. It remains unclear how to reconcile such theories with neural networks, which learn via incremental weight updates and are thus a natural model for the latter type of learning, but are not obviously compatible with the former. However, recent evidence suggests that metalearning neural networks and large language models are capable of "in-context learning" (ICL) -- the ability to flexibly grasp the structure of a new task from a few examples. Here, we show that the dynamic interplay between ICL and default in-weight learning (IWL) naturally captures a broad range of learning phenomena observed in humans, reproducing curriculum effects on category-learning and compositional tasks, and recapitulating a tradeoff between flexibility and retention. Our work shows how emergent ICL can equip neural networks with fundamentally different learning properties that can coexist with their native IWL, thus offering a novel perspective on dual-process theories and human cognitive flexibility.

📄 PDF Abstract BibTeX arXiv:2402.08674

Code (1)

jlrussin/icl-iwl-interplay 공식 구현 pytorch

Tasks

BlockingIn-Context Learning

Similar Papers 제목 키워드 기반

Persuasion at Play: Understanding Misinformation Dynamics in Demographic-Aware Human-LLM Interactions

2025-03-03 · Angana Borah, Rada Mihalcea, Verónica Pérez-Rosas

Existing challenges in misinformation exposure and susceptibility vary across demographic groups, as some populations are more vulnerable to misinformation than others. Large language models (LLMs) introduce new dimensio…

Misinformation

Extending Causal Models from Machines into Humans

2019-10-31 · Severin Kacianka, Amjad Ibrahim, Alexander Pretschner, Alexander Trende 외

Causal Models are increasingly suggested as a means to reason about the behavior of cyber-physical systems in socio-technical contexts. They allow us to analyze courses of events and reason about possible alternatives. U…

Supporting Data-Frame Dynamics in AI-assisted Decision Making

2025-04-22 · Chengbo Zheng, Tim Miller, Alina Bialkowski, H Peter Soyer 외

High stakes decision-making often requires a continuous interplay between evolving evidence and shifting hypotheses, a dynamic that is not well supported by current AI decision support systems. In this paper, we introduc…

Decision MakingDiagnostic

Thorns and Algorithms: Navigating Generative AI Challenges Inspired by Giraffes and Acacias

2024-07-16 · Waqar Hussain

The interplay between humans and Generative AI (Gen AI) draws an insightful parallel with the dynamic relationship between giraffes and acacias on the African Savannah. Just as giraffes navigate the acacia's thorny defen…

MisinformationNavigate

MM-FusionNet: Context-Aware Dynamic Fusion for Multi-modal Fake News Detection with Large Vision-Language Models

2025-08-05 · Junhao He, Tianyu Liu, Jingyuan Zhao, Benjamin Turner arxiv

The proliferation of multi-modal fake news on social media poses a significant threat to public trust and social stability. Traditional detection methods, primarily text-based, often fall short due to the deceptive inter…

Fake News Detection