paper-with-me

Papers

A large scale annotated child language construction database

2012-05-01 · LREC 2012 5 · Aline Villavicencio, Beracah Yankama, Marco Idiart, Robert Berwick

Large scale annotated corpora of child language can be of great value in assessing theoretical proposals regarding language acquisition models. For example, they can help determine whether the type and amount of data required by a proposed language acquisition model can actually be found in a naturalistic data sample. To this end, several recent efforts have augmented the CHILDES child language corpora with POS tagging and parsing information for languages such as English. With the increasing availability of robust NLP systems and electronic resources, these corpora can be further annotated with more detailed information about the properties of words, verb argument structure, and sentences. This paper describes such an initiative for combining information from various sources to extend the annotation of the English CHILDES corpora with linguistic, psycholinguistic and distributional information, along with an example illustrating an application of this approach to the extraction of verb alternation information. The end result, the English CHILDES Verb Construction Database, is an integrated resource containing information such as grammatical relations, verb semantic classes, and age of acquisition, enabling more targeted complex searches involving different levels of annotation that can facilitate a more detailed analysis of the linguistic input available to children.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Language AcquisitionPOSPOS Tagging

Similar Papers 제목 키워드 기반

Multiword Expressions in Child Language

2016-05-01 · LREC 2016 5 · Rodrigo Wilkens, Marco Idiart, Aline Villavicencio

The goal of this work is to introduce CHILDES-MWE, which contains English CHILDES corpora automatically annotated with Multiword Expressions (MWEs) information. The result is a resource with almost 350,000 sentences anno…

Language Acquisition

What Exactly do Children Receive in Language Acquisition? A Case Study on CHILDES with Automated Detection of Filler-Gap Dependencies

2026-03-02 · Zhenghao Herbert Zhou, William Dai, Maya Viswanathan, Simon Charlow 외 arxiv

Children's acquisition of filler-gap dependencies has been argued by some to depend on innate grammatical knowledge, while others suggest that the distributional evidence available in child-directed speech suffices. Unfo…

Language AcquisitionDependency Parsing

ChiSense-12: An English Sense-Annotated Child-Directed Speech Corpus

2022-06-01 · LREC 2022 6 · Francesco Cabiddu, Lewis Bott, Gary Jones, Chiara Gambi

Language acquisition research has benefitted from the use of annotated corpora of child-directed speech to examine key questions about how children learn and process language in real-world contexts. However, a lack of se…

Language AcquisitionWord Sense Disambiguation

Automatic Annotation of Grammaticality in Child-Caregiver Conversations

2024-03-21 · Mitja Nikolaus, Abhishek Agrawal, Petros Kaklamanis, Alex Warstadt 외

The acquisition of grammar has been a central question to adjudicate between theories of language acquisition. In order to conduct faster, more reproducible, and larger-scale corpus studies on grammaticality in child-car…

Language Acquisition

CAIT: A Syntactic Parsing Toolkit for Child-Adult InTeractions

2026-05-19 · Francesca Padovani, Xiulin Yang, Bastian Bunzeck, Jaap Jumelet 외 arxiv

CHILDES is a paramount resource for language acquisition studies -- yet computational tools for analyzing its syntactic structure remain limited. Leveraging the recent release of the UD-English-CHILDES treebank with gold…

Language Acquisition