paper-with-me

홈 › Papers

A Metric for MLLM Alignment in Large-scale Recommendation

2025-08-07 · Yubin Zhang, Yanhua Huang, Haiming Xu, Mingliang Qi, Chang Wang, Jiarui Jin, Xiangyuan Ren, Xiaodan Wang, Ruiwen Xu arxiv

Multimodal recommendation has emerged as a critical technique in modern recommender systems, leveraging content representations from advanced multimodal large language models (MLLMs). To ensure these representations are well-adapted, alignment with the recommender system is essential. However, evaluating the alignment of MLLMs for recommendation presents significant challenges due to three key issues: (1) static benchmarks are inaccurate because of the dynamism in real-world applications, (2) evaluations with online system, while accurate, are prohibitively expensive at scale, and (3) conventional metrics fail to provide actionable insights when learned representations underperform. To address these challenges, we propose the Leakage Impact Score (LIS), a novel metric for multimodal recommendation. Rather than directly assessing MLLMs, LIS efficiently measures the upper bound of preference data. We also share practical insights on deploying MLLMs with LIS in real-world scenarios. Online A/B tests on both Content Feed and Display Ads of Xiaohongshu's Explore Feed production demonstrate the effectiveness of our proposed method, showing significant improvements in user spent time and advertiser value.

📄 PDF Abstract BibTeX arXiv:2508.04963

Code (0)

등록된 구현이 없습니다.

Tasks

Multimodal Recommendation

Similar Papers 제목 키워드 기반

Benchmarking Multimodal Large Language Models for Missing Modality Completion in Product Catalogues

2026-01-27 · Junchen Fu, Wenhao Deng, Kaiwen Zheng, Ioannis Arapakis 외 arxiv

Missing-modality information on e-commerce platforms, such as absent product images or textual descriptions, often arises from annotation errors or incomplete metadata, impairing both product presentation and downstream …

Recommendation Systems

Serendipitous Recommendation with Multimodal LLM

2025-06-09 · HaoTing Wang, Jianling Wang, Hao Li, Fangjun Yi 외

Conventional recommendation systems succeed in identifying relevant content but often fail to provide users with surprising or novel items. Multimodal Large Language Models (MLLMs) possess the world knowledge and multimo…

Recommendation SystemsWorld Knowledge

Evaluating Cognitive Age Alignment in Interactive AI Agents

2026-05-18 · Yifan Shen, Jiawen Zhang, Jian Xu, Junho Kim 외 arxiv

While agentic AI and its core multimodal large language models (MLLMs) have demonstrated remarkable promise in language and visual reasoning across domains ranging from daily life to advanced scientific research, a profo…

Visual Reasoning

GeoAlign: Geometric Feature Realignment for MLLM Spatial Reasoning

2026-04-14 · Zhaochen Liu, Limeng Qiao, Guanglu Wan, Tingting Jiang arxiv

Multimodal large language models (MLLMs) have exhibited remarkable performance in various visual tasks, yet still struggle with spatial reasoning. Recent efforts mitigate this by injecting geometric features from 3D foun…

Spatial Reasoning

Harnessing Multimodal Large Language Models for Multimodal Sequential Recommendation

2024-08-19 · Yuyang Ye, Zhi Zheng, Yishan Shen, Tianshu Wang 외

Recent advances in Large Language Models (LLMs) have demonstrated significant potential in the field of Recommendation Systems (RSs). Most existing studies have focused on converting user behavior logs into textual promp…

Large Language ModelMultimodal Large Language ModelMulti-modal RecommendationMultimodal Recommendation+2