paper-with-me

홈 › Papers

Controlling Language Difficulty in Dialogues with Linguistic Features

2025-09-18 · Shuyao Xu, Wenguang Wang, Handong Gao, Wei Kang, Long Qin, Weizhi Wang arxiv

Large language models (LLMs) have emerged as powerful tools for supporting second language acquisition, particularly in simulating interactive dialogues for speaking practice. However, adapting the language difficulty of LLM-generated responses to match learners' proficiency levels remains a challenge. This work addresses this issue by proposing a framework for controlling language proficiency in educational dialogue systems. Our approach leverages three categories of linguistic features, readability features (e.g., Flesch-Kincaid Grade Level), syntactic features (e.g., syntactic tree depth), and lexical features (e.g., simple word ratio), to quantify and regulate text complexity. We demonstrate that training LLMs on linguistically annotated dialogue data enables precise modulation of language proficiency, outperforming prompt-based methods in both flexibility and stability. To evaluate this, we introduce Dilaprix, a novel metric integrating the aforementioned features, which shows strong correlation with expert judgments of language difficulty. Empirical results reveal that our approach achieves superior controllability of language proficiency while maintaining high dialogue quality.

📄 PDF Abstract BibTeX arXiv:2509.14545

Code (0)

등록된 구현이 없습니다.

Tasks

Language Acquisition

Similar Papers 제목 키워드 기반

Linguistic Cues to Deception and Perceived Deception in Interview Dialogues

2018-06-01 · NAACL 2018 6 · Sarah Ita Levitan, Angel Maredia, Julia Hirschberg

We explore deception detection in interview dialogues. We analyze a set of linguistic features in both truthful and deceptive responses to interview questions. We also study the perception of deception, identifying chara…

BIG-bench Machine LearningDeception DetectionGeneral Classification

Discourse Structure and Dialogue Acts in Multiparty Dialogue: the STAC Corpus

2016-05-01 · LREC 2016 5 · Nicholas Asher, Julie Hunter, Mathieu Morey, Benamara Farah 외

This paper describes the STAC resource, a corpus of multi-party chats annotated for discourse structure in the style of SDRT (Asher and Lascarides, 2003; Lascarides and Asher, 2009). The main goal of the STAC project is …

Annotation and Detection of Emotion in Text-based Dialogue Systems with CNN

2017-10-03 · Jialiang Zhao, Qi Gao

Knowledge of users' emotion states helps improve human-computer interaction. In this work, we presented EmoNet, an emotion detector of Chinese daily dialogues based on deep convolutional neural networks. In order to main…

A Linguistic Analysis of Visually Grounded Dialogues Based on Spatial Expressions

2020-10-07 · Findings of the Association for Computational Linguistics 2020 · Takuma Udagawa, Takato Yamazaki, Akiko Aizawa

Recent models achieve promising results in visually grounded dialogues. However, existing datasets often contain undesirable biases and lack sophisticated linguistic analyses, which make it difficult to understand how we…

Coreference ResolutionNatural Language Visual GroundingSpatial Relation Recognition

Understanding Linguistic Accommodation in Code-Switched Human-Machine Dialogues

2020-11-01 · CONLL 2020 · Tanmay Parekh, Emily Ahn, Yulia Tsvetkov, Alan W Black

Code-switching is a ubiquitous phenomenon in multilingual communities. Natural language technologies that wish to communicate like humans must therefore adaptively incorporate code-switching techniques when they are depl…