paper-with-me

홈 › Papers

CLIMB: Curriculum Learning for Infant-inspired Model Building

2023-11-15 · Richard Diehl Martinez, Zebulon Goriely, Hope McGovern, Christopher Davis, Andrew Caines, Paula Buttery, Lisa Beinborn

We describe our team's contribution to the STRICT-SMALL track of the BabyLM Challenge. The challenge requires training a language model from scratch using only a relatively small training dataset of ten million words. We experiment with three variants of cognitively-motivated curriculum learning and analyze their effect on the performance of the model on linguistic evaluation tasks. In the vocabulary curriculum, we analyze methods for constraining the vocabulary in the early stages of training to simulate cognitively more plausible learning curves. In the data curriculum experiments, we vary the order of the training instances based on i) infant-inspired expectations and ii) the learning behavior of the model. In the objective curriculum, we explore different variations of combining the conventional masked language modeling task with a more coarse-grained word class prediction task to reinforce linguistic generalization capabilities. Our results did not yield consistent improvements over our own non-curriculum learning baseline across a range of linguistic benchmarks; however, we do find marginal gains on select tasks. Our analysis highlights key takeaways for specific combinations of tasks and settings which benefit from our proposed curricula. We moreover determine that careful selection of model architecture, and training hyper-parameters yield substantial improvements over the default baselines provided by the BabyLM challenge.

📄 PDF Abstract BibTeX arXiv:2311.08886

Code (0)

등록된 구현이 없습니다.

Tasks

Language ModelingLanguage ModellingMasked Language Modelingmodel

Similar Papers 제목 키워드 기반

CLIMB: Language-Guided Continual Learning for Task Planning with Iterative Model Building

2024-10-17 · Walker Byrnes, Miroslav Bogdanovic, Avi Balakirsky, Stephen Balakirsky 외

Intelligent and reliable task planning is a core capability for generalized robotics, requiring a descriptive domain representation that sufficiently models all object and state information for the scene. We present CLIM…

Continual LearningDescriptiveRobot Task PlanningTask Planning

Baby Sophia: A Developmental Approach to Self-Exploration through Self-Touch and Hand Regard

2025-11-12 · Stelios Zarifis, Ioannis Chalkiadakis, Artemis Chardouveli, Vasiliki Moutzouri 외 arxiv

Inspired by infant development, we propose a Reinforcement Learning (RL) framework for autonomous self-exploration in a robotic agent, Baby Sophia, using the BabyBench simulation environment. The agent learns self-touch …

Reinforcement Learning

Curriculum Learning With Infant Egocentric Videos

2023-09-21 · NeurIPS 2023 11

Infants possess a remarkable ability to rapidly learn and process visual inputs. As an infant's mobility increases, so does the variety and dynamics of their visual inputs. Is this change in the properties of the visual …

Reinforcement Learning-based Robust Wall Climbing Locomotion Controller in Ferromagnetic Environment

2025-10-23 · Yong Um, Young-Ha Shin, Joon-Ha Kim, Soonpyo Kwon 외 arxiv

We present a reinforcement learning framework for quadrupedal wall-climbing locomotion that explicitly addresses uncertainty in magnetic foot adhesion. A physics-based adhesion model of a quadrupedal magnetic climbing ro…

Reinforcement Learning

Learning Transferability: A Two-Stage Reinforcement Learning Approach for Enhancing Quadruped Robots' Performance in U-Shaped Stair Climbing

2026-02-16 · Baixiao Huang, Baiyu Huang, Yu Hou arxiv

Quadruped robots are employed in various scenarios in building construction. However, autonomous stair climbing across different indoor staircases remains a major challenge for robot dogs to complete building constructio…

Reinforcement Learning