paper-with-me

Papers Native Language Identification

“Native Language Identification” 태그가 달린 논문 88편 · 필터 해제

The Impact of Editorial Intervention on Detecting Native Language Traces

2026-05-11 · Ahmet Yavuz Uluslu, Mark Gales, Kate Knill, Gerold Schneider arxiv

Native Language Identification (NLI) is the task of determining an author's native language (L1) from their non-native writing. With the advent of human-AI co-authorship, learner texts are routinely corrected and rewritt…

Native Language IdentificationGrammatical Error Correction

Can We Still Hear the Accent? Investigating the Resilience of Native Language Signals in the LLM Era

2026-03-20 · Nabelanita Utami, Ryohei Sasano arxiv

The evolution of writing assistance tools from machine translation to large language models (LLMs) has changed how researchers write. This study investigates whether this shift is homogenizing research papers by analyzin…

Native Language IdentificationMachine Translation

I can tell whether you are a Native Hawlêri Speaker! How ANN, CNN, and RNN perform in NLI-Native Language Identification

2026-02-11 · Hardi Garari, Hossein Hassani arxiv

Native Language Identification (NLI) is a task in Natural Language Processing (NLP) that typically determines the native language of an author through their writing or a speaker through their speaking. It has various app…

Native Language Identification

NLP Privacy Risk Identification in Social Media (NLP-PRISM): A Survey

2026-01-26 · Dhiman Goswami, Jai Kruthunz Naveen Kumar, Sanchari Das arxiv

Natural Language Processing (NLP) is integral to social media analytics but often processes content containing Personally Identifiable Information (PII), behavioral cues, and metadata raising privacy risks such as survei…

Native Language IdentificationSentiment Analysis

Robust Native Language Identification through Agentic Decomposition

2025-09-20 · Ahmet Yavuz Uluslu, Tannon Kew, Tilia Ellendorff, Gerold Schneider 외 arxiv

Large language models (LLMs) often achieve high performance in native language identification (NLI) benchmarks by leveraging superficial contextual clues such as names, locations, and cultural stereotypes, rather than th…

Native Language Identification

Leveraging Open-Source Large Language Models for Native Language Identification

2024-09-15 · Yee Man Ng, Ilia Markov

Native Language Identification (NLI) - the task of identifying the native language (L1) of a person based on their writing in the second language (L2) - has applications in forensics, marketing, and second language acqui…

Feature EngineeringLanguage AcquisitionLanguage IdentificationMarketing+2

Native Language Identification with Large Language Models

2023-12-13 · Wei zhang, Alexandre Salle

We present the first experiments on Native Language Identification (NLI) using LLMs such as GPT-4. NLI is the task of predicting a writer's first language by analyzing their writings in a second language, and is used in …

Language AcquisitionLanguage IdentificationNative Language Identification

Native Language Identification with Big Bird Embeddings

2023-09-13 · Sergey Kramp, Giovanni Cassani, Chris Emmery

Native Language Identification (NLI) intends to classify an author's native language based on their writing in another language. Historically, the task has heavily relied on time-consuming linguistic feature engineering,…

Computational EfficiencyFeature EngineeringLanguage IdentificationNative Language Identification

Turkish Native Language Identification

2023-07-27 · Ahmet Yavuz Uluslu, Gerold Schneider

In this paper, we present the first application of Native Language Identification (NLI) for the Turkish language. NLI involves predicting the writer's first language by analysing their writing in different languages. Whi…

Language IdentificationNative Language Identification

Scaling Native Language Identification with Transformer Adapters

2022-11-18 · Ahmet Yavuz Uluslu, Gerold Schneider

Native language identification (NLI) is the task of automatically identifying the native language (L1) of an individual based on their language production in a learned language. It is useful for a variety of purposes inc…

Language IdentificationMarketingMulti-Label ClassificationMUlTI-LABEL-ClASSIFICATION+1

Unravelling Interlanguage Facts via Explainable Machine Learning

2022-08-02 · Barbara Berti, Andrea Esuli, Fabrizio Sebastiani

Native language identification (NLI) is the task of training (via supervised machine learning) a classifier that guesses the native language of the author of a text. This task has been extensively researched in the last …

BIG-bench Machine LearningLanguage IdentificationNative Language Identification

Native Language Identification and Reconstruction of Native Language Relationship Using Japanese Learner Corpus

2021-11-01 · PACLIC 2021 11 · Mitsuhiro Nishijima, Ying Liu
Language IdentificationNative Language Identification

A Deep Generative Approach to Native Language Identification

2020-12-01 · COLING 2020 8 · Ehsan Lotfi, Ilia Markov, Walter Daelemans

Native language identification (NLI) {--} identifying the native language (L1) of a person based on his/her writing in the second language (L2) {--} is useful for a variety of purposes, including marketing, security, and…

BIG-bench Machine LearningLanguage IdentificationLanguage ModellingMarketing+2

Native-Language Identification with Attention

2020-12-01 · ICON 2020 12 · Stian Steinbakken, Björn Gambäck

The paper explores how an attention-based approach can increase performance on the task of native-language identification (NLI), i.e., to identify an author’s first language given information expressed in a second langua…

Language IdentificationNative Language Identification

Investigating the effect of auxiliary objectives for the automated grading of learner English speech transcriptions

2020-07-01 · ACL 2020 6 · Hannah Craighead, Andrew Caines, Paula Buttery, Helen Yannakoudakis

We address the task of automatically grading the language proficiency of spontaneous speech based on textual features from automatic speech recognition transcripts. Motivated by recent advances in multi-task learning, we…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Language IdentificationLanguage Modeling+5

A Report on the 2020 VUA and TOEFL Metaphor Detection Shared Task

2020-07-01 · WS 2020 7 · Chee Wee (Ben) Leong, Beata Beigman Klebanov, Chris Hamill, Egon Stemle 외

In this paper, we report on the shared task on metaphor identification on VU Amsterdam Metaphor Corpus and on a subset of the TOEFL Native Language Identification Corpus. The shared task was conducted as apart of the ACL…

Language IdentificationNative Language Identification

Topics to Avoid: Demoting Latent Confounds in Text Classification

2019-09-01 · IJCNLP 2019 11 · Sachin Kumar, Shuly Wintner, Noah A. Smith, Yulia Tsvetkov

Despite impressive performance on many text classification tasks, deep neural networks tend to learn frequent superficial patterns that are specific to the training data and do not always generalize well. In this work, w…

ClassificationGeneral ClassificationLanguage IdentificationNative Language Identification+2

Towards Ethical Content-Based Detection of Online Influence Campaigns

2019-08-29 · Evan Crothers, Nathalie Japkowicz, Herna Viktor

The detection of clandestine efforts to influence users in online communities is a challenging problem with significant active development. We demonstrate that features derived from the text of user comments are useful f…

Language IdentificationNative Language IdentificationSentence

Regression or classification? Automated Essay Scoring for Norwegian

2019-08-01 · WS 2019 8 · Stig Johan Berggren, Taraka Rama, Lilja {\O}vrelid

In this paper we present first results for the task of Automated Essay Scoring for Norwegian learner language. We analyze a number of properties of this task experimentally and assess (i) the formulation of the task as e…

Automated Essay ScoringBIG-bench Machine LearningClassificationGeneral Classification+4

Anglicized Words and Misspelled Cognates in Native Language Identification

2019-08-01 · WS 2019 8 · Ilia Markov, Vivi Nastase, Carlo Strapparava

In this paper, we present experiments that estimate the impact of specific lexical choices of people writing in a second language (L2). In particular, we look at misspelled words that indicate lexical uncertainty on the …

Language IdentificationNative Language Identification
1–20 / 88 다음 →