Resources for Turkish Natural Language Processing: A critical survey
This paper presents a comprehensive survey of corpora and lexical resources available for Turkish. We review a broad range of resources, focusing on the ones that are publicly available. In addition to providing information about the available linguistic resources, we present a set of recommendations, and identify gaps in the data available for conducting research and building applications in Turkish Linguistics and Natural Language Processing.
Code (0)
등록된 구현이 없습니다.
Tasks
SurveySimilar Papers 제목 키워드 기반
Building Foundations for Natural Language Processing of Historical Turkish: Resources and Models
This paper introduces foundational resources and models for natural language processing (NLP) of historical Turkish, a domain that has remained underexplored in computational linguistics. We present the first named entit…
Dependency ParsingDomain Adaptationnamed-entity-recognitionNamed Entity Recognition+3Fine-tuning Transformer-based Encoder for Turkish Language Understanding Tasks
Deep learning-based and lately Transformer-based language models have been dominating the studies of natural language processing in the last years. Thanks to their accurate and fast fine-tuning characteristics, they have…
named-entity-recognitionNamed Entity RecognitionNatural Language UnderstandingQuestion Answering+4Resources for Turkish Dependency Parsing: Introducing the BOUN Treebank and the BoAT Annotation Tool
In this paper, we introduce the resources that we developed for Turkish dependency parsing, which include a novel manually annotated treebank (BOUN Treebank), along with the guidelines we adopted, and a new annotation to…
ArticlesCultural Vocal Bursts Intensity PredictionDependency ParsingChallenges Encountered in Turkish Natural Language Processing Studies
Natural language processing is a branch of computer science that combines artificial intelligence with linguistics. It aims to analyze a language element such as writing or speaking with software and convert it into info…
DiversityTUR2SQL: A Cross-Domain Turkish Dataset For Text-to-SQL
The field of converting natural language into corresponding SQL queries using deep learning techniques has attracted significant attention in recent years. While existing Text-to-SQL datasets primarily focus on English a…
Text to SQLText-To-SQL