paper-with-me

홈 › Papers

SmurfCat at SemEval-2024 Task 6: Leveraging Synthetic Data for Hallucination Detection

2024-04-09 · Elisei Rykov, Yana Shishkina, Kseniia Petrushina, Kseniia Titova, Sergey Petrakov, Alexander Panchenko

In this paper, we present our novel systems developed for the SemEval-2024 hallucination detection task. Our investigation spans a range of strategies to compare model predictions with reference standards, encompassing diverse baselines, the refinement of pre-trained encoders through supervised learning, and an ensemble approaches utilizing several high-performing models. Through these explorations, we introduce three distinct methods that exhibit strong performance metrics. To amplify our training data, we generate additional training samples from unlabelled training subset. Furthermore, we provide a detailed comparative analysis of our approaches. Notably, our premier method achieved a commendable 9th place in the competition's model-agnostic track and 17th place in model-aware track, highlighting its effectiveness and potential.

📄 PDF Abstract BibTeX arXiv:2404.06137

Code (1)

s-nlp/shroom 공식 구현 pytorch

Tasks

Hallucination

Similar Papers 제목 키워드 기반

SmurfCat at PAN 2024 TextDetox: Alignment of Multilingual Transformers for Text Detoxification

2024-07-07 · Elisei Rykov, Konstantin Zaytsev, Ivan Anisimov, Alexandr Voronin

This paper presents a solution for the Multilingual Text Detoxification task in the PAN-2024 competition of the SmurfCat team. Using data augmentation through machine translation and a special filtering procedure, we col…

Data AugmentationMachine Translation

AmazUtah_NLP at SemEval-2024 Task 9: A MultiChoice Question Answering System for Commonsense Defying Reasoning

2024-05-16 · Mina Ghashami, Soumya Smruti Mishra

The SemEval 2024 BRAINTEASER task represents a pioneering venture in Natural Language Processing (NLP) by focusing on lateral thinking, a dimension of cognitive reasoning that is often overlooked in traditional linguisti…

Multiple-choiceQuestion AnsweringSentence

MALTO at SemEval-2024 Task 6: Leveraging Synthetic Data for LLM Hallucination Detection

2024-03-01 · Federico Borra, Claudio Savelli, Giacomo Rosso, Alkis Koudounas 외

In Natural Language Generation (NLG), contemporary Large Language Models (LLMs) face several challenges, such as generating fluent yet inaccurate outputs and reliance on fluency-centric metrics. This often leads to neura…

Data AugmentationHallucinationNatural Language InferenceSentence+1

DCU-SEManiacs at SemEval-2016 Task 1: Synthetic Paragram Embeddings for Semantic Textual Similarity

2016-06-01 · SEMEVAL 2016 6 · Chris Hokamp, Piyush Arora
Machine TranslationSemantic Textual SimilaritySentence Embeddings

SemEval-2025 Task 4: Unlearning sensitive content from Large Language Models

2025-04-02 · Anil Ramakrishna, Yixin Wan, Xiaomeng Jin, Kai-Wei Chang 외

We introduce SemEval-2025 Task 4: unlearning sensitive content from Large Language Models (LLMs). The task features 3 subtasks for LLM unlearning spanning different use cases: (1) unlearn long form synthetic creative doc…

Form