paper-with-me

Papers

Enhancing Large Language Model Induced Task-Oriented Dialogue Systems Through Look-Forward Motivated Goals

2023-09-16 · Zhiyuan Hu, Yue Feng, Yang Deng, Zekun Li, See-Kiong Ng, Anh Tuan Luu, Bryan Hooi

Recently, the development of large language models (LLMs) has been significantly enhanced the question answering and dialogue generation, and makes them become increasingly popular in current practical scenarios. While unlike the general dialogue system which emphasizes the semantic performance, the task-oriented dialogue (ToD) systems aim to achieve the dialogue goal efficiently and successfully in multiple turns. Unfortunately, existing LLM-induced ToD systems lack the direct reward toward the final goal and do not take account of the dialogue proactivity that can strengthen the dialogue efficiency. To fill these gaps, we introduce the ProToD (Proactively Goal-Driven LLM-Induced ToD) approach, which anticipates the future dialogue actions and incorporates the goal-oriented reward signal to enhance ToD systems. Additionally, we present a novel evaluation method that assesses ToD systems based on goal-driven dialogue simulations. This method allows us to gauge user satisfaction, system efficiency and successful rate while overcoming the limitations of current Information and Success metrics. Empirical experiments conducted on the MultiWoZ 2.1 dataset demonstrate that our model can achieve superior performance using only 10% of the data compared to previous end-to-end fully supervised models. This improvement is accompanied by enhanced user satisfaction and efficiency.

📄 PDF Abstract BibTeX arXiv:2309.08949

Code (0)

등록된 구현이 없습니다.

Tasks

Dialogue GenerationLanguage ModelingLanguage ModellingLarge Language ModelQuestion AnsweringTask-Oriented Dialogue Systems

Similar Papers 제목 키워드 기반

Does syntax matter? A strong baseline for Aspect-based Sentiment Analysis with RoBERTa

2021-04-11 · NAACL 2021 4 · Junqi Dai, Hang Yan, Tianxiang Sun, PengFei Liu 외

Aspect-based Sentiment Analysis (ABSA), aiming at predicting the polarities for aspects, is a fine-grained task in the field of sentiment analysis. Previous work showed syntactic information, e.g. dependency trees, can e…

Aspect-Based Sentiment AnalysisAspect-Based Sentiment Analysis (ABSA)Dependency ParsingSentiment Analysis

The Hallucination Dilemma: Factuality-Aware Reinforcement Learning for Large Reasoning Models

2025-05-30 · Junyi Li, Hwee Tou Ng

Large language models (LLMs) have significantly advanced in reasoning tasks through reinforcement learning (RL) optimization, achieving impressive capabilities across various challenging benchmarks. However, our empirica…

HallucinationMathematical ReasoningReinforcement Learning (RL)

SPS: Steering Probability Squeezing for Better Exploration in Reinforcement Learning for Large Language Models

2026-04-18 · Yifu Huo, Chenglong Wang, Ziming Zhu, Shunjie Xing 외 arxiv

Reinforcement learning (RL) has emerged as a promising paradigm for training reasoning-oriented models by leveraging rule-based reward signals. However, RL training typically tends to improve single-sample success rates …

Reinforcement Learning

Zero-shot Cross-lingual Dialogue Systems with Transferable Latent Variables

2019-11-11 · IJCNLP 2019 11 · Zihan Liu, Jamin Shin, Yan Xu, Genta Indra Winata 외

Despite the surging demands for multilingual task-oriented dialog systems (e.g., Alexa, Google Home), there has been less research done in multilingual or cross-lingual scenarios. Hence, we propose a zero-shot adaptation…

Intent DetectionNatural Language Understandingslot-fillingSlot Filling

Harnessing Large Language Models for Intelligent Resource Allocation in the Internet of Everything

2026-07-29 · Haijun Zhang, Zhuojun Duan, Zijun Wu, Xu Ma 외 arxiv

The rapid development of the Internet of Everything (IoE) is accelerating the adoption of intelligent applications. However, the massive number of connected devices generates diverse and heterogeneous tasks, which pose i…