HeLI-based Experiments in Discriminating Between Dutch and Flemish Subtitles
This paper presents the experiments and results obtained by the SUKI team in the Discriminating between Dutch and Flemish in Subtitles shared task of the VarDial 2018 Evaluation Campaign. Our best submission was ranked 8th, obtaining macro F1-score of 0.61. Our best results were produced by a language identifier implementing the HeLI method without any modifications. We describe, in addition to the best method we used, some of the experiments we did with unsupervised clustering.
Code (0)
등록된 구현이 없습니다.
Tasks
ClusteringLanguage IdentificationText CategorizationSimilar Papers 제목 키워드 기반
STEVENDU2018's system in VarDial 2018: Discriminating between Dutch and Flemish in Subtitles
This paper introduces the submitted system for team STEVENDU2018 during VarDial 2018 Discriminating between Dutch and Flemish in Subtitles(DFS). Post evaluation analyses are also presented, the results obtained indicate …
When Simple n-gram Models Outperform Syntactic Approaches: Discriminating between Dutch and Flemish
In this paper we present the results of our participation in the Discriminating between Dutch and Flemish in Subtitles VarDial 2018 shared task. We try techniques proven to work well for discriminating between language v…
Exploring Classifier Combinations for Language Variety Identification
This paper describes CLiPS{'}s submissions for the Discriminating between Dutch and Flemish in Subtitles (DFS) shared task at VarDial 2018. We explore different ways to combine classifiers trained on different feature gr…
Language IdentificationPOSClassifier Ensembles for Dialect and Language Variety Identification
In this paper we present ensemble-based systems for dialect and language variety identification using the datasets made available by the organizers of the VarDial Evaluation Campaign 2018. We present a system developed t…
Dialect IdentificationIdentification of Differences between Dutch Language Varieties with the VarDial2018 Dutch-Flemish Subtitle Data
With the goal of discovering differences between Belgian and Netherlandic Dutch, we participated as Team Taurus in the Dutch-Flemish Subtitles task of VarDial2018. We used a rather simple marker-based method, but a wide …
PositionText Classification