paper-with-me

홈 › Papers

Controllable Spoken Dialogue Generation: An LLM-Driven Grading System for K-12 Non-Native English Learners

2026-04-24 · Haidong Yuan, Haokun Zhao, Wanshi Xu, Songjun Cao, Qingyu Zhou, Long Ma, Hongjie Fan arxiv

Large language models (LLMs) often fail to meet the pedagogical needs of K-12 English learners in non-native contexts due to a proficiency mismatch. To address this widespread challenge, we introduce a proficiency-aligned framework that adapts LLM outputs to learner abilities, using China's national curriculum (CSE) as a representative case. Our framework enables precise control over lexical complexity through a four-tier grading system, supported by a comprehensive suite of new resources: graded vocabulary lists and a multi-turn dialogue corpus. Our core technical contribution is the \textbf{DDPO} algorithm,Diversity Driven Policy Optimization, a multi-turn GRPO-based approach designed to preserve dialogue diversity while holistically optimizing dialogue quality. This method significantly outperforms conventional approaches, achieving low out-of-vocabulary rates and high diversity while enhancing conversational naturalness and pedagogical value. While grounded in the CSE, our framework is designed for flexibility and can be readily adapted to other educational standards. Our models, data, and code will all be open-sourced, providing a scalable platform for personalized English speaking practice that effectively addresses the unique challenges faced by K-12 learners in non-immersive environments.

📄 PDF Abstract BibTeX arXiv:2604.22542

Code (0)

등록된 구현이 없습니다.

Tasks

Dialogue Generation

Similar Papers 제목 키워드 기반

TiCo: Time-Controllable Spoken Dialogue Model

2026-03-23 · Kai-Wei Chang, Wei-Chih Chen, En-Pei Hu, Hung-yi Lee 외 arxiv

We introduce TiCo, a time-controllable spoken dialogue model (SDM) that follows time-constrained instructions (e.g., "Please generate a response lasting about 15 seconds") and generates spoken responses with controllable…

Reinforcement LearningInstruction Following

UltraVoice: Scaling Fine-Grained Style-Controlled Speech Conversations for Spoken Dialogue Models

2025-10-26 · Wenming Tu, Guanrou Yang, Ruiqi Yan, Wenxi Chen 외 arxiv

Spoken dialogue models currently lack the ability for fine-grained speech style control, a critical capability for human-like interaction that is often overlooked in favor of purely functional capabilities like reasoning…

Instruction FollowingQuestion AnsweringSpeech Synthesis

Commonsense-Aware Prompting for Controllable Empathetic Dialogue Generation

2023-02-02 · Yiren Liu, Halil Kilicoglu

Improving the emotional awareness of pre-trained language models is an emerging important problem for dialogue generation tasks. Although prior studies have introduced methods to improve empathetic dialogue generation, f…

Dialogue Generation

SLAM-Omni: Timbre-Controllable Voice Interaction System with Single-Stage Training

2024-12-20 · Wenxi Chen, Ziyang Ma, Ruiqi Yan, Yuzhe Liang 외

Recent advancements highlight the potential of end-to-end real-time spoken dialogue systems, showcasing their low latency and high quality. In this paper, we introduce SLAM-Omni, a timbre-controllable, end-to-end voice i…

Spoken Dialogue Systems

ZipVoice-Dialog: Non-Autoregressive Spoken Dialogue Generation with Flow Matching

2025-07-12 · Han Zhu, Wei Kang, Liyong Guo, Zengwei Yao 외

Generating spoken dialogue is more challenging than monologue text-to-speech (TTS) due to the need for realistic turn-taking and distinct speaker timbres. Existing spoken dialogue generation models, being auto-regressive…

Dialogue Generationtext-to-speechText to Speech