ETPC - A Paraphrase Identification Corpus Annotated with Extended Paraphrase Typology and Negation
Code (1)
Tasks
Natural Language InferenceNegationParaphrase IdentificationQuestion AnsweringSemantic Textual SimilarityText SimplificationText SummarizationSimilar Papers 제목 키워드 기반
A Corpus of Word-Aligned Asked and Anticipated Questions in a Virtual Patient Dialogue System
We present a corpus of virtual patient dialogues to which we have added manually annotated gold standard word alignments. Since each question asked by a medical student in the dialogues is mapped to a canonical, anticipa…
BnPC: A Corpus for Paraphrase Detection in Bangla
In this paper, we present the first benchmark dataset for paraphrase detection in Bangla language. Despite being the sixth most spoken language in the world, paraphrase identification in the Bangla language is barely ex…
Paraphrase IdentificationSentenceFinnish Paraphrase Corpus
In this paper, we introduce the first fully manually annotated paraphrase corpus for Finnish containing 53,572 paraphrase pairs harvested from alternative subtitles and news headings. Out of all paraphrase pairs in our c…
AllMahaParaphrase: A Marathi Paraphrase Detection Corpus and BERT-based Models
Paraphrases are a vital tool to assist language understanding tasks such as question answering, style transfer, semantic parsing, and data augmentation tasks. Indic languages are complex in natural language processing (N…
Question AnsweringData AugmentationSemantic ParsingStyle TransferAutomatic Compilation of Resources for Academic Writing and Evaluating with Informal Word Identification and Paraphrasing System
We present the first approach to automatically building resources for academic writing. The aim is to build a writing aid system that automatically edits a text so that it better adheres to the academic style of writing.…
Paraphrase Generation