paper-with-me

Papers

References Indeed Matter? Reference-Free Preference Optimization for Conversational Query Reformulation

2025-05-10 · Doyoung Kim, YoungJun Lee, Joeun Kim, Jihwan Bang, Hwanjun Song, Susik Yoon, Jae-Gil Lee

Conversational query reformulation (CQR) has become indispensable for improving retrieval in dialogue-based applications. However, existing approaches typically rely on reference passages for optimization, which are impractical to acquire in real-world scenarios. To address this limitation, we introduce a novel reference-free preference optimization framework DualReform that generates pseudo reference passages from commonly-encountered conversational datasets containing only queries and responses. DualReform attains this goal through two key innovations: (1) response-based inference, where responses serve as proxies to infer pseudo reference passages, and (2) response refinement via the dual-role of CQR, where a CQR model refines responses based on the shared objectives between response refinement and CQR. Despite not relying on reference passages, DualReform achieves 96.9--99.1% of the retrieval accuracy attainable only with reference passages and surpasses the state-of-the-art method by up to 31.6%.

📄 PDF Abstract BibTeX arXiv:2505.06552

Code (0)

등록된 구현이 없습니다.

Tasks

Retrieval

Similar Papers 제목 키워드 기반

Eliciting Worker Preference for Task Completion

2018-01-10 · Mohammadreza Esfandiari, Senjuti Basu Roy, Sihem Amer-Yahia

Current crowdsourcing platforms provide little support for worker feedback. Workers are sometimes invited to post free text describing their experience and preferences in completing tasks. They can also use forums such a…

Cold-Start Personalization via Training-Free Priors from Structured World Models

2026-02-16 · Avinandan Bose, Shuyue Stella Li, Faeze Brahman, Pang Wei Koh 외 arxiv

Cold-start personalization requires inferring user preferences through interaction when no user-specific historical data is available. The core challenge is a routing problem: each task admits dozens of preference dimens…

Reinforcement LearningBayesian Inference

Collaborative filtering to capture AI user's preferences as norms

2023-08-01 · Marc Serramia, Natalia Criado, Michael Luck

Customising AI technologies to each user's preferences is fundamental to them functioning well. Unfortunately, current methods require too much user involvement and fail to capture their true preferences. In fact, to avo…

Collaborative FilteringRecommendation Systems

Individual characteristics associated with risk and time preferences: A multi country representative survey

2022-04-28 · Thomas Meissner, Xavier Gassmann, Corinne Faure, Joachim Schleich

This paper empirically analyzes how individual characteristics are associated with risk aversion, loss aversion, time discounting, and present bias. To this end, we conduct a large-scale demographically representative su…

Inferring Lexicographically-Ordered Rewards from Preferences

2022-02-21 · Alihan Hüyük, William R. Zame, Mihaela van der Schaar

Modeling the preferences of agents over a set of alternatives is a principal concern in many areas. The dominant approach has been to find a single reward/utility function with the property that alternatives yielding hig…