paper-with-me

Papers

LecEval: An Automated Metric for Multimodal Knowledge Acquisition in Multimedia Learning

2025-05-04 · Joy Lim Jia Yin, Daniel Zhang-li, Jifan Yu, Haoxuan Li, Shangqing Tu, Yuanchun Wang, Zhiyuan Liu, Huiqin Liu, Lei Hou, Juanzi Li, Bin Xu

Evaluating the quality of slide-based multimedia instruction is challenging. Existing methods like manual assessment, reference-based metrics, and large language model evaluators face limitations in scalability, context capture, or bias. In this paper, we introduce LecEval, an automated metric grounded in Mayer's Cognitive Theory of Multimedia Learning, to evaluate multimodal knowledge acquisition in slide-based learning. LecEval assesses effectiveness using four rubrics: Content Relevance (CR), Expressive Clarity (EC), Logical Structure (LS), and Audience Engagement (AE). We curate a large-scale dataset of over 2,000 slides from more than 50 online course videos, annotated with fine-grained human ratings across these rubrics. A model trained on this dataset demonstrates superior accuracy and adaptability compared to existing metrics, bridging the gap between automated and human assessments. We release our dataset and toolkits at https://github.com/JoylimJY/LecEval.

📄 PDF Abstract BibTeX arXiv:2505.02078

Code (1)

joylimjy/leceval 공식 구현 pytorch

Tasks

Language ModelingLanguage ModellingLarge Language Model

Similar Papers 제목 키워드 기반

DyMRL: Dynamic Multispace Representation Learning for Multimodal Event Forecasting in Knowledge Graph

2026-03-25 · Feng Zhao, Kangzheng Liu, Teng Peng, Yu Yang 외 arxiv

Accurate representation of multimodal knowledge is crucial for event forecasting in real-world scenarios. However, existing studies have largely focused on static settings, overlooking the dynamic acquisition and fusion …

Representation LearningLogical Reasoning

Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos

2025-01-23 · Kairui Hu, Penghao Wu, Fanyi Pu, Wang Xiao 외

Humans acquire knowledge through three cognitive stages: perceiving information, comprehending knowledge, and adapting knowledge to solve novel problems. Videos serve as an effective medium for this learning process, fac…

BiosecurID: a multimodal biometric database

2021-11-02 · Julian Fierrez, Javier Galbally, Javier Ortega-Garcia, Manuel R Freire 외

A new multimodal biometric database, acquired in the framework of the BiosecurID project, is presented together with the description of the acquisition setup and protocol. The database includes eight unimodal biometric t…

CLBench-V: Evaluating Multimodal Context Learning from Grounding to Knowledge Acquisition

2026-07-28 · Lai Wei, Chengqi Li, Jiapeng Li, Ruina Hu 외 arxiv

Real-world tasks often require models to learn from task-specific context rather than relying only on pre-trained knowledge. While recent work has highlighted this capability as context learning, existing evaluations mai…

Visual Question AnsweringSpatial Reasoning

Benchmarking Multimodal Knowledge Conflict for Large Multimodal Models

2025-05-26 · Yifan Jia, Kailin Jiang, Yuyang Liang, Qihan Ren 외

Large Multimodal Models(LMMs) face notable challenges when encountering multimodal knowledge conflicts, particularly under retrieval-augmented generation(RAG) frameworks where the contextual information from external sou…

BenchmarkingRAGRetrieval-augmented Generation