paper-with-me

홈 › Papers

Contribuer au progr\`es solidaire des recherches et de la documentation : la Collection Pangloss et la Collection AuCo (Contributing to joint progress in documentation and research: some achievements and future perspectives of the Pangloss Collection and the AuCo Collection)

2016-07-01 · JEPTALNRECITAL 2016 7 · Alexis Michaud, S{\'e}verine Guillaume, Guillaume Jacques, {\DJ}{\u{a}}ng-Khoa Mạc, Michel Jacobson, Thu-H{\`a} Phạm, Matthew Deo

La pr{\'e}sente communication pr{\'e}sente les projets scientifiques et les r{\'e}alisations de deux collections h{\'e}berg{\'e}es par la plateforme de ressources orales Cocoon : la Collection Pangloss, qui concerne principalement des langues de tradition orale (sans {\'e}criture), du monde entier ; et la Collection AuCo, d{\'e}di{\'e}e aux langues du Vietnam et de pays voisins. L{'}objectif est un progr{\`e}s solidaire des recherches et de la documentation linguistique. L{'}accent est mis sur les perspectives ouvertes pour la recherche en phon{\'e}tique/phonologie par certaines r{\'e}alisations r{\'e}centes dans le cadre de ces deux Collections.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Learning Semantic Correspondences in Technical Documentation

2017-05-13 · ACL 2017 7 · Kyle Richardson, Jonas Kuhn

We consider the problem of translating high-level textual descriptions to formal representations in technical documentation as part of an effort to model the meaning of such documentation. We focus specifically on the pr…

Semantic Parsing

Building a Time-Aligned Cross-Linguistic Reference Corpus from Language Documentation Data (DoReCo)

2020-05-01 · LREC 2020 5 · Ludger Paschen, Fran{\c{c}}ois Delafontaine, Christoph Draxler, Susanne Fuchs 외

Natural speech data on many languages have been collected by language documentation projects aiming to preserve lingustic and cultural traditions in audivisual records. These data hold great potential for large-scale cro…

Documenting Geographically and Contextually Diverse Data Sources: The BigScience Catalogue of Language Data and Resources

2022-01-25 · Angelina McMillan-Major, Zaid Alyafeai, Stella Biderman, Kimbo Chen 외

In recent years, large-scale data collection efforts have prioritized the amount of data collected in order to improve the modeling capabilities of large language models. This prioritization, however, has resulted in con…

TEDI: Trustworthy and Ethical Dataset Indicators to Analyze and Compare Dataset Documentation

2025-05-23 · Wiebke Hutiri, Mircea Cimpoi, Morgan Scheuerman, Victoria Matthews 외

Dataset transparency is a key enabler of responsible AI, but insights into multimodal dataset attributes that impact trustworthy and ethical aspects of AI applications remain scarce and are difficult to compare across da…

Outiller une langue peu dot\'ee gr\^ace au TALN : l'exemple du corse et de la BDLC (Tooling up a less-resourced language with NLP : the example of Corsican and BDLC)

2019-07-01 · JEPTALNRECITAL 2019 7 · Laurent Kevers, Florian Gu{\'e}niot, Aurelia Ghjacumina Tognotti, Stella Retali-Medori

Nos recherches sur la langue corse nous am{\`e}nent naturellement {\`a} envisager l{'}utilisation d{'}outils pour le traitement automatique du langage. Apr{\`e}s une br{\`e}ve introduction sur le corse et sur le projet q…