Paraphrase Identification
11개 벤치마크 · 논문 174편 · 이 태스크의 논문 보기 →
Benchmarks
Quora Question Pairs
MSRP
Quora Question Pairs Dev
2017_test set
AP
IMDb
PIT
TURL
WikiHop
Yelp
Most implemented
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
XLNet: Generalized Autoregressive Pretraining for Language Understanding
data2vec: A General Framework for Self-supervised Learning in Speech, Vision and Language
FNet: Mixing Tokens with Fourier Transforms
TinyBERT: Distilling BERT for Natural Language Understanding
Bilateral Multi-Perspective Matching for Natural Language Sentences
Papers
LeWiDi-2025 at NLPerspectives: Third Edition of the Learning with Disagreements Shared Task
Many researchers have reached the conclusion that AI models should be trained to be aware of the possibility of variation and disagreement in human judgments, and evaluated as per their ability to recognize such variatio…
Natural Language InferenceParaphrase IdentificationSarcasm DetectionCompressed Models are NOT Trust-equivalent to Their Large Counterparts
Large Deep Learning models are often compressed before being deployed in a resource-constrained environment. Can we trust the prediction of compressed models just as we trust the prediction of the original large model? E…
Natural Language InferenceParaphrase IdentificationText ClassificationEvaluating the Effectiveness of Linguistic Knowledge in Pretrained Language Models: A Case Study of Universal Dependencies
Universal Dependencies (UD), while widely regarded as the most successful linguistic framework for cross-lingual syntactic representation, remains underexplored in terms of its effectiveness. This paper addresses this ga…
Paraphrase IdentificationEnhancing Paraphrase Type Generation: The Impact of DPO and RLHF Evaluated with Human-Ranked Data
Paraphrasing re-expresses meaning to enhance applications like text simplification, machine translation, and question-answering. Specific paraphrase types facilitate accurate semantic analysis and robust language models.…
Machine TranslationParaphrase GenerationParaphrase IdentificationQuestion Answering+2Enhancing Plagiarism Detection in Marathi with a Weighted Ensemble of TF-IDF and BERT Embeddings for Low-Resource Language Processing
Plagiarism involves using another person's work or concepts without proper attribution, presenting them as original creations. With the growing amount of data communicated in regional languages such as Marathi -- one of …
Paraphrase IdentificationSentenceSentence EmbeddingsTougher Text, Smarter Models: Raising the Bar for Adversarial Defence Benchmarks
Recent advancements in natural language processing have highlighted the vulnerability of deep learning models to adversarial attacks. While various defence mechanisms have been proposed, there is a lack of comprehensive …
Adversarial RobustnessBenchmarkingNatural Language InferenceParaphrase Identification+2