paper-with-me

홈 › Papers

X-Reflect: Cross-Reflection Prompting for Multimodal Recommendation

2024-08-27 · Hanjia Lyu, Ryan Rossi, Xiang Chen, Md Mehrab Tanjim, Stefano Petrangeli, Somdeb Sarkhel, Jiebo Luo

Large Language Models (LLMs) and Large Multimodal Models (LMMs) have been shown to enhance the effectiveness of enriching item descriptions, thereby improving the accuracy of recommendation systems. However, most existing approaches either rely on text-only prompting or employ basic multimodal strategies that do not fully exploit the complementary information available from both textual and visual modalities. This paper introduces a novel framework, Cross-Reflection Prompting, termed X-Reflect, designed to address these limitations by prompting LMMs to explicitly identify and reconcile supportive and conflicting information between text and images. By capturing nuanced insights from both modalities, this approach generates more comprehensive and contextually richer item representations. Extensive experiments conducted on two widely used benchmarks demonstrate that our method outperforms existing prompting baselines in downstream recommendation accuracy. Additionally, we evaluate the generalizability of our framework across different LMM backbones and the robustness of the prompting strategies, offering insights for optimization. This work underscores the importance of integrating multimodal information and presents a novel solution for improving item understanding in multimodal recommendation systems.

📄 PDF Abstract BibTeX arXiv:2408.15172

Code (0)

등록된 구현이 없습니다.

Tasks

Multimodal RecommendationRecommendation Systems

Similar Papers 제목 키워드 기반

SRPO: Enhancing Multimodal LLM Reasoning via Reflection-Aware Reinforcement Learning

2025-06-02 · Zhongwei Wan, Zhihao Dou, Che Liu, Yu Zhang 외

Multimodal large language models (MLLMs) have shown promising capabilities in reasoning tasks, yet still struggle with complex problems requiring explicit self-reflection and self-correction, especially compared to their…

Multimodal Reasoningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Can Large Language Models Self-Correct in Medical Question Answering? An Exploratory Study

2026-03-31 · Zaifu Zhan, Mengyuan Cui, Rui Zhang arxiv

Large language models (LLMs) have achieved strong performance on medical question answering (medical QA), and chain-of-thought (CoT) prompting has further improved results by eliciting explicit intermediate reasoning; me…

Question Answering

Unveiling the Latent Directions of Reflection in Large Language Models

2025-08-23 · Fu-Chieh Chang, Yu-Ting Lee, Pei-Yuan Wu arxiv

Reflection, the ability of large language models (LLMs) to evaluate and revise their own reasoning, has been widely used to improve performance on complex reasoning tasks. Yet, most prior works emphasizes designing refle…

Reinforcement Learning

Enhancing Sequential Recommendations through Multi-Perspective Reflections and Iteration

2024-09-10 · Weicong Qin, Yi Xu, Weijie Yu, Chenglei Shen 외

Sequence recommendation (SeqRec) aims to predict the next item a user will interact with by understanding user intentions and leveraging collaborative filtering information. Large language models (LLMs) have shown great …

Collaborative FilteringGPU

Reflective Translation: Improving Low-Resource Machine Translation via Structured Self-Reflection

2026-01-27 · Nicholas Cheng arxiv

Low-resource languages such as isiZulu and isiXhosa face persistent challenges in machine translation due to limited parallel data and linguistic resources. Recent advances in large language models suggest that self-refl…

Machine Translation