Discriminative Phrase Embedding for Paraphrase Identification
This work, concerning paraphrase identification task, on one hand contributes to expanding deep learning embeddings to include continuous and discontinuous linguistic phrases. On the other hand, it comes up with a new scheme TF-KLD-KNN to learn the discriminative weights of words and phrases specific to paraphrase task, so that a weighted sum of embeddings can represent sentences more effectively. Based on these two innovations we get competitive state-of-the-art performance on paraphrase identification.
Code (0)
등록된 구현이 없습니다.
Tasks
Paraphrase IdentificationSimilar Papers 제목 키워드 기반
Cross-lingual paraphrase identification
The paraphrase identification task involves measuring semantic similarity between two short sentences. It is a tricky task, and multilingual paraphrase identification is even more challenging. In this work, we train a bi…
Cross-Lingual Paraphrase IdentificationParaphrase IdentificationSemantic SimilaritySemantic Textual SimilarityBetter Early than Late: Fusing Topics with Word Embeddings for Neural Question Paraphrase Identification
Question paraphrase identification is a key task in Community Question Answering (CQA) to determine if an incoming question has been previously asked. Many current models use word embeddings to identify duplicate questio…
Community Question AnsweringParaphrase IdentificationQuestion AnsweringTopic Models+1Towards Automatic Short Answer Assessment for Finnish as a Paraphrase Retrieval Task
Automatic grouping of textual answers has the potential of allowing batch grading, but is challenging because the answers, especially longer essays, have many claims. To explore the feasibility of grouping together answe…
Paraphrase IdentificationRetrievalSentenceSentence EmbeddingsImproving Large-scale Paraphrase Acquisition and Generation
This paper addresses the quality issues in existing Twitter-based paraphrase datasets, and discusses the necessity of using two separate definitions of paraphrase for identification and generation tasks. We present a new…
Language ModelingLanguage ModellingParaphrase GenerationParaphrase Identification+1RuPAWS: A Russian Adversarial Dataset for Paraphrase Identification
Paraphrase identification task can be easily challenged by changing word order, e.g. as in “Can a good person become bad?”. While for English this problem was tackled by the PAWS dataset (Zhang et al., 2019), datasets fo…
Paraphrase Identification