paper-with-me

홈 › Papers

Where and What: Reasoning Dynamic and Implicit Preferences in Situated Conversational Recommendation

2026-04-22 · Dongding Lin, Jian Wang, Yongqi Li, Wenjie Li arxiv

Situated conversational recommendation (SCR), which utilizes visual scenes grounded in specific environments and natural language dialogue to deliver contextually appropriate recommendations, has emerged as a promising research direction due to its close alignment with real-world scenarios. Compared to traditional recommendations, SCR requires a deeper understanding of dynamic and implicit user preferences, as the surrounding scene often influences users' underlying interests, while both may evolve across conversations. This complexity significantly impacts the timing and relevance of recommendations. To address this, we propose situated preference reasoning (SiPeR), a novel framework that integrates two core mechanisms: (1) Scene transition estimation, which estimates whether the current scene satisfies user needs, and guides the user toward a more suitable scene when necessary; and (2) Bayesian inverse inference, which leverages the likelihood of multimodal large language models (MLLMs) to predict user preferences about candidate items within the scene. Extensive experiments on two representative benchmarks demonstrate SiPeR's superiority in both recommendation accuracy and response generation quality. The code and data are available at https://github.com/DongdingLin/SiPeR.

📄 PDF Abstract BibTeX arXiv:2604.20749

Code (0)

등록된 구현이 없습니다.

Tasks

Response Generation

Similar Papers 제목 키워드 기반

TUR-DPO: Topology- and Uncertainty-Aware Direct Preference Optimization

2026-04-30 · Abdulhady Abas Abdullah, Fatemeh Daneshfar, Seyedali Mirjalili, Mourad Oussalah arxiv

Aligning large language models (LLMs) with human preferences is commonly done via reinforcement learning from human feedback (RLHF) with Proximal Policy Optimization (PPO) or, more simply, via Direct Preference Optimizat…

Reinforcement LearningMathematical ReasoningQuestion Answering

The Morality of Probability: How Implicit Moral Biases in LLMs May Shape the Future of Human-AI Symbiosis

2025-09-12 · Eoin O'Doherty, Nicole Weinrauch, Andrew Talone, Uri Klempner 외 arxiv

Artificial intelligence (AI) is advancing at a pace that raises urgent questions about how to align machine decision-making with human moral values. This working paper investigates how leading AI systems prioritize moral…

Preferences Implicit in the State of the World

2019-02-12 · ICLR 2019 5 · Rohin Shah, Dmitrii Krasheninnikov, Jordan Alexander, Pieter Abbeel 외

Reinforcement learning (RL) agents optimize only the features specified in a reward function and are indifferent to anything left out inadvertently. This means that we must not only specify what to do, but also the much …

Reinforcement LearningReinforcement Learning (RL)

A Knowledge Driven Approach to Adaptive Assistance Using Preference Reasoning and Explanation

2020-12-05 · Jason R. Wilson, Leilani Gilpin, Irina Rabkina

There is a need for socially assistive robots (SARs) to provide transparency in their behavior by explaining their reasoning. Additionally, the reasoning and explanation should represent the user's preferences and goals.…

An Extension-based Approach for Computing and Verifying Preferences in Abstract Argumentation

2024-03-26 · Quratul-ain Mahesar, Nir Oren, Wamberto W. Vasconcelos

We present an extension-based approach for computing and verifying preferences in an abstract argumentation system. Although numerous argumentation semantics have been developed previously for identifying acceptable sets…

Abstract Argumentation