paper-with-me

Papers

ProductAgent: Benchmarking Conversational Product Search Agent with Asking Clarification Questions

2024-07-01 · Jingheng Ye, Yong Jiang, Xiaobin Wang, Yinghui Li, Yangning Li, Hai-Tao Zheng, Pengjun Xie, Fei Huang

This paper introduces the task of product demand clarification within an e-commercial scenario, where the user commences the conversation with ambiguous queries and the task-oriented agent is designed to achieve more accurate and tailored product searching by asking clarification questions. To address this task, we propose ProductAgent, a conversational information seeking agent equipped with abilities of strategic clarification question generation and dynamic product retrieval. Specifically, we develop the agent with strategies for product feature summarization, query generation, and product retrieval. Furthermore, we propose the benchmark called PROCLARE to evaluate the agent's performance both automatically and qualitatively with the aid of a LLM-driven user simulator. Experiments show that ProductAgent interacts positively with the user and enhances retrieval performance with increasing dialogue turns, where user demands become gradually more explicit and detailed. All the source codes will be released after the review anonymity period.

📄 PDF Abstract BibTeX arXiv:2407.00942

Code (0)

등록된 구현이 없습니다.

Tasks

BenchmarkingQuestion GenerationQuestion-GenerationRetrieval

Similar Papers 제목 키워드 기반

MTR-Suite: A Framework for Evaluating and Synthesizing Conversational Retrieval Benchmarks

2026-05-20 · Junhao Ruan, Abudukeyumu Abudula, Bei Li, Yongjing Yin 외 arxiv

Accurate evaluation of conversational retrieval is pivotal for advancing Retrieval-Augmented Generation (RAG) systems. However, existing conversational retrieval benchmarks suffer from costly, sparse human annotation or …

DuCCAE: A Hybrid Engine for Immersive Conversation via Collaboration, Augmentation, and Evolution

2026-02-25 · Xin Shen, Zhishu Jiang, Jiaye Yang, Haibo Liu 외 arxiv

Immersive conversational systems in production face a persistent trade-off between responsiveness and long-horizon task capability. Real-time interaction is achievable for lightweight turns, but requests involving planni…

Response Generation

Dynamic benchmarking framework for LLM-based conversational data capture

2025-02-04 · Pietro Alessandro Aluffi, Patrick Zietkiewicz, Marya Bazzi, Matt Arderne 외

The rapid evolution of large language models (LLMs) has transformed conversational agents, enabling complex human-machine interactions. However, evaluation frameworks often focus on single tasks, failing to capture the d…

Benchmarking

MEMTRACK: Evaluating Long-Term Memory and State Tracking in Multi-Platform Dynamic Agent Environments

2025-10-01 · Darshan Deshpande, Varun Gangal, Hersh Mehta, Anand Kannappan 외 arxiv

Recent works on context and memory benchmarking have primarily focused on conversational instances but the need for evaluating memory in dynamic enterprise environments is crucial for its effective application. We introd…

Conversational Recommendation System with Unsupervised Learning

2016-09-22 · Yueming Sun, Yi Zhang, Yunfei Chen, Roger Jin

We will demonstrate a conversational products recommendation agent. This system shows how we combine research in personalized recommendation systems with research in dialogue systems to build a virtual sales agent. Based…

Conversational RecommendationRecommendation Systems