paper-with-me

홈 › Papers

DialogueForge: LLM Simulation of Human-Chatbot Dialogue

2025-07-21 · Ruizhe Zhu, Hao Zhu, Yaxuan Li, Syang Zhou, Shijing Cai, Malgorzata Lazuka, Elliott Ash arxiv

Collecting human-chatbot dialogues typically demands substantial manual effort and is time-consuming, which limits and poses challenges for research on conversational AI. In this work, we propose DialogueForge - a framework for generating AI-simulated conversations in human-chatbot style. To initialize each generated conversation, DialogueForge uses seed prompts extracted from real human-chatbot interactions. We test a variety of LLMs to simulate the human chatbot user, ranging from state-of-the-art proprietary models to small-scale open-source LLMs, and generate multi-turn dialogues tailored to specific tasks. In addition, we explore fine-tuning techniques to enhance the ability of smaller models to produce indistinguishable human-like dialogues. We evaluate the quality of the simulated conversations and compare different models using the UniEval and GTEval evaluation protocols. Our experiments show that large proprietary models (e.g., GPT-4o) generally outperform others in generating more realistic dialogues, while smaller open-source models (e.g., Llama, Mistral) offer promising performance with greater customization. We demonstrate that the performance of smaller models can be significantly improved by employing supervised fine-tuning techniques. Nevertheless, maintaining coherent and natural long-form human-like dialogues remains a common challenge across all models.

📄 PDF Abstract BibTeX arXiv:2507.15752

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

LLM Roleplay: Simulating Human-Chatbot Interaction

2024-07-04 · Hovhannes Tamoyan, Hendrik Schuff, Iryna Gurevych

The development of chatbots requires collecting a large number of human-chatbot dialogues to reflect the breadth of users' sociodemographic backgrounds and conversational goals. However, the resource requirements to cond…

Chatbot

Synthetic Users, Real Differences: an Evaluation Framework for User Simulation in Multi-Turn Conversations

2026-05-04 · Yu Lu Liu, Hyokun Yun, Tanya Roosta, Ziang Xiao arxiv

There is growing interest in exploring user simulation as an alternative to gathering and scoring real user-chatbot interactions for AI chatbot evaluation. For this purpose, it is important to ensure the realism of the s…

Ensemble-Based Deep Reinforcement Learning for Chatbots

2019-08-27 · Heriberto Cuayáhuitl, Donghyeon Lee, Seonghan Ryu, Yongjin Cho 외

Trainable chatbots that exhibit fluent and human-like conversations remain a big challenge in artificial intelligence. Deep Reinforcement Learning (DRL) is promising for addressing this challenge, but its successful appl…

ChatbotClusteringDeep Reinforcement LearningOpen-Ended Question Answering+4

Positively transitioned sentiment dialogue corpus for developing emotion-affective open-domain chatbots

2022-08-09 · Weixuan Wang, Wei Peng, Chong Hsuan Huang, Haoran Wang

In this paper, we describe a data enhancement method for developing Emily, an emotion-affective open-domain chatbot. The proposed method is based on explicitly modeling positively transitioned (PT) sentiment data from mu…

Chatbot

The Open-domain Paradox for Chatbots: Common Ground as the Basis for Human-like Dialogue

2023-03-21 · Gabriel Skantze, A. Seza Doğruöz

There is a surge in interest in the development of open-domain chatbots, driven by the recent advancements of large language models. The "openness" of the dialogue is expected to be maximized by providing minimal informa…

Position