paper-with-me

홈 › Papers

BAMBI: Developing Baby Language Models for Italian

2025-03-12 · Alice Suozzi, Luca Capone, Gianluca E. Lebani, Alessandro Lenci

This paper presents BAMBI (BAby language Models Boostrapped for Italian), a series of Baby Language Models (BabyLMs) trained on data that mimic the linguistic input received by a five-year-old Italian-speaking child. The BAMBI models are tested using a benchmark specifically designed to evaluate language models, which takes into account the amount of training input the models received. The BAMBI models are compared against a large language model (LLM) and a multimodal language model (VLM) to study the contribution of extralinguistic information for language acquisition. The results of our evaluation align with the existing literature on English language models, confirming that while reduced training data support the development of relatively robust syntactic competence, they are insufficient for fostering semantic understanding. However, the gap between the training resources (data and computation) of the BAMBI models and the LLMs is not fully reflected in their performance: despite LLMs' massive training, their performance is not much better than that of BAMBI models. This suggests that strategies beyond scaling training resources, such as data curation, inclusion of multimodal input, and other training strategies such as curriculum learning, could play a crucial role in shaping model performance.

📄 PDF Abstract BibTeX arXiv:2503.09481

Code (0)

등록된 구현이 없습니다.

Tasks

Language AcquisitionLanguage ModelingLanguage ModellingLarge Language Model

Methods 이 논문이 사용한 방법론

ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

BAMBINO-LM: (Bilingual-)Human-Inspired Continual Pretraining of BabyLM

2024-06-17 · Zhewen Shen, Aditya Joshi, Ruey-Cheng Chen

Children from bilingual backgrounds benefit from interactions with parents and teachers to re-acquire their heritage language. In this paper, we investigate how this insight from behavioral study can be incorporated into…

Continual Pretrainingzero-shot-classificationZero-Shot Learning

BabyStories: Can Reinforcement Learning Teach Baby Language Models to Write Better Stories?

2023-10-25 · Xingmeng Zhao, Tongnian Wang, Sheri Osborn, Anthony Rios

Language models have seen significant growth in the size of their corpus, leading to notable performance improvements. Yet, there has been limited progress in developing models that handle smaller, more human-like datase…

BAMBI: blind accelerated multimodal Bayesian inference

2011-10-13 · Philip Graff, Farhan Feroz, Michael P. Hobson, Anthony Lasenby

In this paper we present an algorithm for rapid Bayesian analysis that combines the benefits of nested sampling and artificial neural networks. The blind accelerated multimodal Bayesian inference (BAMBI) algorithm implem…

Bayesian Inference

LLM-BABYBENCH: Understanding and Evaluating Grounded Planning and Reasoning in LLMs

2025-05-17 · Omar Choukrani, Idriss Malek, Daniil Orel, Zhuohan Xie 외

Assessing the capacity of Large Language Models (LLMs) to plan and reason within the constraints of interactive environments is crucial for developing capable AI agents. We introduce $\textbf{LLM-BabyBench}$, a new bench…

Task 2

CTAP for Italian: Integrating Components for the Analysis of Italian into a Multilingual Linguistic Complexity Analysis Tool

2020-05-01 · LREC 2020 5 · Nadezda Okinina, Jennifer-Carmen Frey, Zarah Weiss

Linguistic complexity research being a very actively developing field, an increasing number of text analysis tools are created that use natural language processing techniques for the automatic extraction of quantifiable …

Diversity