paper-with-me

홈 › Papers

Overview of AMALGUM – Large Silver Quality Annotations across English Genres

2021-02-01 · SCiL 2021 2 · Luke Gessler, Siyao Peng, Yang Liu, YIlun Zhu, Shabnam Behzad, Amir Zeldes
📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Midas Loop: A Prioritized Human-in-the-Loop Annotation for Large Scale Multilayer Data

2022-06-01 · LREC (LAW) 2022 6 · Luke Gessler, Lauren Levine, Amir Zeldes

Large scale annotation of rich multilayer corpus data is expensive and time consuming, motivating approaches that integrate high quality automatic tools with active learning in order to prioritize human labeling of hard …

Active LearningManagementSegmentationSentence+1

AMALGUM -- A Free, Balanced, Multilayer English Web Corpus

2020-06-18 · LREC 2020 5 · Luke Gessler, Siyao Peng, Yang Liu, YIlun Zhu 외

We present a freely available, genre-balanced English web corpus totaling 4M tokens and featuring a large number of high-quality automatic annotation layers, including dependency trees, non-named entity annotations, core…

coreference-resolutionCoreference Resolution

NAVCON: A Cognitively Inspired and Linguistically Grounded Corpus for Vision and Language Navigation

2024-12-17 · Karan Wanchoo, Xiaoye Zuo, Hannah Gonzalez, Soham Dan 외

We present NAVCON, a large-scale annotated Vision-Language Navigation (VLN) corpus built on top of two popular datasets (R2R and RxR). The paper introduces four core, cognitively motivated and linguistically grounded, na…

Few-Shot LearningVision and Language NavigationVision-Language Navigation

Smelting Gold and Silver for Improved Multilingual AMR-to-Text Generation

2021-09-08 · EMNLP 2021 11 · Leonardo F. R. Ribeiro, Jonas Pfeiffer, Yue Zhang, Iryna Gurevych

Recent work on multilingual AMR-to-text generation has exclusively focused on data augmentation strategies that utilize silver AMR. However, this assumes a high quality of generated AMRs, potentially limiting the transfe…

AMR-to-Text GenerationData AugmentationText Generation

A Probabilistic Annotation Model for Crowdsourcing Coreference

2018-10-01 · EMNLP 2018 10 · Silviu Paun, Jon Chamberlain, Udo Kruschwitz, Juntao Yu 외

The availability of large scale annotated corpora for coreference is essential to the development of the field. However, creating resources at the required scale via expert annotation would be too expensive. Crowdsourcin…

Coreference ResolutionmodelQuestion Answering