paper-with-me

홈 › Papers

RefCo and its Checker: Improving Language Documentation Corpora’s Reusability Through a Semi-Automatic Review Process

2022-06-01 · LREC 2022 6 · Herbert Lange, Jocelyn Aznar

The QUEST (QUality ESTablished) project aims at ensuring the reusability of audio-visual datasets (Wamprechtshammer et al., 2022) by devising quality criteria and curating processes. RefCo (Reference Corpora) is an initiative within QUEST in collaboration with DoReCo (Documentation Reference Corpus, Paschen et al. (2020)) focusing on language documentation projects. Previously, Aznar and Seifart (2020) introduced a set of quality criteria dedicated to documenting fieldwork corpora. Based on these criteria, we establish a semi-automatic review process for existing and work-in-progress corpora, in particular for language documentation. The goal is to improve the quality of a corpus by increasing its reusability. A central part of this process is a template for machine-readable corpus documentation and automatic data verification based on this documentation. In addition to the documentation and automatic verification, the process involves a human review and potentially results in a RefCo certification of the corpus. For each of these steps, we provide guidelines and manuals. We describe the evaluation process in detail, highlight the current limits for automatic evaluation and how the manual review is organized accordingly.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Toward Reusability of AI Models Using Dynamic Updates of AI Documentation

2026-04-19 · Peter Bajcsy, Walid Keyrouz arxiv

This work addresses the challenge of disseminating reusable artificial intelligence (AI) models accompanied by AI documentation (a.k.a., AI model cards). The work is motivated by the large number of trained AI models tha…

Modeling Context in Referring Expressions

2016-07-31 · Licheng Yu, Patrick Poirson, Shan Yang, Alexander C. Berg 외

Humans refer to objects in their environments all the time, especially in dialogue with other people. We explore generating and comprehending natural language referring expressions for objects in images. In particular, w…

Referring ExpressionReferring expression generationText Generation

Neuro-Symbolic AI for LEED compliance: Document-Centric Benchmarking, Deterministic Numeric Checking, and When Multimodal Hurts

2026-07-17 · Aritro De, Juliana Felkner arxiv

LEED v4.1 BD+C certification remains a document-intensive process that requires reviewers to read hundreds of pages of project evidence and apply credit-specific threshold logic by hand. This paper investigates whether s…

Towards Unified Referring Expression Segmentation Across Omni-Level Visual Target Granularities

2025-04-02 · Jing Liu, Wenxuan Wang, Yisi Zhang, Yepeng Tang 외

Referring expression segmentation (RES) aims at segmenting the entities' masks that match the descriptive language expression. While traditional RES methods primarily address object-level grounding, real-world scenarios …

DescriptiveLarge Language ModelMultimodal Large Language ModelObject+3

Auto-BenchmarkCard: Automated Synthesis of Benchmark Documentation

2025-12-10 · Aris Hofmann, Inge Vejsbjerg, Dhaval Salwala, Elizabeth M. Daly arxiv

We present Auto-BenchmarkCard, a workflow for generating validated descriptions of AI benchmarks. Benchmark documentation is often incomplete or inconsistent, making it difficult to interpret and compare benchmarks acros…