SpeechTrans@SMM4H’20: Impact of Preprocessing and N-grams on Automatic Classification of Tweets That Mention Medications
This paper describes our system developed for automatically classifying tweets that mention medications. We used the Decision Tree classifier for this task. We have shown that using some elementary preprocessing steps and TF-IDF n-grams led to acceptable classifier performance. Indeed, the F1-score recorded was 74.58% in the development phase and 63.70% in the test phase.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
The Speechtransformer for Large-scale Mandarin Chinese Speech Recognition
Attention-based sequence-to-sequence architectures have made great progress in the speech recognition task. The SpeechTransformer, a no-recurrence encoder-decoder architecture, has shown promising results on small-scale …
Decoderspeech-recognitionSpeech RecognitionTUN: Detecting Significant Points in Persistence Diagrams with Deep Learning
Persistence diagrams (PDs) provide a powerful tool for understanding the topology of the underlying shape of a point cloud. However, identifying which points in PDs encode genuine signals remains challenging. This challe…
Lexical-semantic resources: yet powerful resources for automatic personality classification
In this paper, we aim to reveal the impact of lexical-semantic resources, used in particular for word sense disambiguation and sense-level semantic categorization, on automatic personality classification task. While styl…
ClassificationGeneral ClassificationWord Sense DisambiguationOSACT4 Shared Task on Offensive Language Detection: Intensive Preprocessing-Based Approach
The preprocessing phase is one of the key phases within the text classification pipeline. This study aims at investigating the impact of the preprocessing phase on text classification, specifically on offensive language …
ClassificationDimensionality ReductionGeneral ClassificationHate Speech Detection+2How Document Pre-processing affects Keyphrase Extraction Performance
The SemEval-2010 benchmark dataset has brought renewed attention to the task of automatic keyphrase extraction. This dataset is made up of scientific articles that were automatically converted from PDF format to plain te…
ArticlesKeyphrase Extraction