paper-with-me

홈 › Papers

InterFeedback: Unveiling Interactive Intelligence of Large Multimodal Models via Human Feedback

2025-02-20 · Henry Hengyuan Zhao, Wenqi Pei, Yifei Tao, Haiyang Mei, Mike Zheng Shou

Existing benchmarks do not test Large Multimodal Models (LMMs) on their interactive intelligence with human users, which is vital for developing general-purpose AI assistants. We design InterFeedback, an interactive framework, which can be applied to any LMM and dataset to assess this ability autonomously. On top of this, we introduce InterFeedback-Bench which evaluates interactive intelligence using two representative datasets, MMMU-Pro and MathVerse, to test 10 different open-source LMMs. Additionally, we present InterFeedback-Human, a newly collected dataset of 120 cases designed for manually testing interactive performance in leading models such as OpenAI-o1 and Claude-3.5-Sonnet. Our evaluation results indicate that even the state-of-the-art LMM, OpenAI-o1, struggles to refine its responses based on human feedback, achieving an average score of less than 50%. Our findings point to the need for methods that can enhance LMMs' capabilities to interpret and benefit from feedback.

📄 PDF Abstract BibTeX arXiv:2502.15027

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Towards Interactive Intelligence for Digital Humans

2025-12-15 · Yiyi Cai, Xuangeng Chu, Xiwei Gao, Sitong Gong 외 arxiv

We introduce Interactive Intelligence, a novel paradigm of digital human that is capable of personality-aligned expression, adaptive interaction, and self-evolution. To realize this, we present Mio (Multimodal Interactiv…

COIN: Conversational Interactive Networks for Emotion Recognition in Conversation

2021-06-01 · NAACL (maiworkshop) 2021 6 · Haidong Zhang, Yekun Chai

Emotion recognition in conversation has received considerable attention recently because of its practical industrial applications. Existing methods tend to overlook the immediate mutual interaction between different spea…

Emotion RecognitionEmotion Recognition in Conversation

MEIA: Multimodal Embodied Perception and Interaction in Unknown Environments

2024-02-01 · Yang Liu, Xinshuai Song, Kaixuan Jiang, Weixing Chen 외

With the surge in the development of large language models, embodied intelligence has attracted increasing attention. Nevertheless, prior works on embodied intelligence typically encode scene or historical memory in an u…

Embodied Question AnsweringLanguage ModelingLanguage ModellingLarge Language Model+2

Evaluating Cognitive Age Alignment in Interactive AI Agents

2026-05-18 · Yifan Shen, Jiawen Zhang, Jian Xu, Junho Kim 외 arxiv

While agentic AI and its core multimodal large language models (MLLMs) have demonstrated remarkable promise in language and visual reasoning across domains ranging from daily life to advanced scientific research, a profo…

Visual Reasoning

LifeEval: A Multimodal Benchmark for Assistive AI in Egocentric Daily Life Tasks

2026-02-28 · Hengjian Gao, Kaiwei Zhang, Shibo Wang, Mingjie Chen 외 arxiv

The rapid progress of Multimodal Large Language Models (MLLMs) marks a significant step toward artificial general intelligence, offering great potential for augmenting human capabilities. However, their ability to provid…