paper-with-me

홈 › Papers

Challenges Encountered in Turkish Natural Language Processing Studies

2021-01-21 · Kadir Tohma, Yakup Kutlu

Natural language processing is a branch of computer science that combines artificial intelligence with linguistics. It aims to analyze a language element such as writing or speaking with software and convert it into information. Considering that each language has its own grammatical rules and vocabulary diversity, the complexity of the studies in this field is somewhat understandable. For instance, Turkish is a very interesting language in many ways. Examples of this are agglutinative word structure, consonant/vowel harmony, a large number of productive derivational morphemes (practically infinite vocabulary), derivation and syntactic relations, a complex emphasis on vocabulary and phonological rules. In this study, the interesting features of Turkish in terms of natural language processing are mentioned. In addition, summary info about natural language processing techniques, systems and various sources developed for Turkish are given.

📄 PDF Abstract BibTeX arXiv:2101.11436

Code (0)

등록된 구현이 없습니다.

Tasks

Diversity

Methods 이 논문이 사용한 방법론

INFO This study presents the analysis and principle of an innovative optimizer named weIghted meaN oF vectOrs (INFO) to optimize different problems. INFO is a modified weight mean…

Similar Papers 제목 키워드 기반

Bridging the Bosphorus: Advancing Turkish Large Language Models through Strategies for Low-Resource Language Adaptation and Benchmarking

2024-05-07 · Emre Can Acikgoz, Mete Erdogan, Deniz Yuret

Large Language Models (LLMs) are becoming crucial across various fields, emphasizing the urgency for high-quality models in underrepresented languages. This study explores the unique challenges faced by low-resource lang…

BenchmarkingModel SelectionTransfer Learning

Building Foundations for Natural Language Processing of Historical Turkish: Resources and Models

2025-01-08 · Şaziye Betül Özateş, Tarık Emre Tıraş, Ece Elif Adak, Berat Doğan 외

This paper introduces foundational resources and models for natural language processing (NLP) of historical Turkish, a domain that has remained underexplored in computational linguistics. We present the first named entit…

Dependency ParsingDomain Adaptationnamed-entity-recognitionNamed Entity Recognition+3

Resources for Turkish Natural Language Processing: A critical survey

2022-04-11 · Çağrı Çöltekin, A. Seza Doğruöz, Özlem Çetinoğlu

This paper presents a comprehensive survey of corpora and lexical resources available for Turkish. We review a broad range of resources, focusing on the ones that are publicly available. In addition to providing informat…

Survey

Detecting Code-Switching between Turkish-English Language Pair

2018-11-01 · WS 2018 11 · Zeynep Yirmibe{\c{s}}o{\u{g}}lu, G{\"u}l{\c{s}}en Eryi{\u{g}}it

Code-switching (usage of different languages within a single conversation context in an alternative manner) is a highly increasing phenomenon in social media and colloquial usage which poses different challenges for natu…

Information Retrieval

Doğal Dil İşlemede Tokenizasyon Standartları ve Ölçümü: Türkçe Üzerinden Büyük Dil Modellerinin Karşılaştırmalı Analizi

2025-08-18 · M. Ali Bayram, Ali Arda Fincan, Ahmet Semih Gümüş, Sercan Karakaş 외 arxiv

Tokenization is a fundamental preprocessing step in Natural Language Processing (NLP), significantly impacting the capability of large language models (LLMs) to capture linguistic and semantic nuances. This study introdu…