paper-with-me

Papers

Multimodal Late Fusion Model for Problem-Solving Strategy Classification in a Machine Learning Game

2025-07-30 · Clemens Witt, Thiemo Leonhardt, Nadine Bergner, Mareen Grillenberger arxiv

Machine learning models are widely used to support stealth assessment in digital learning environments. Existing approaches typically rely on abstracted gameplay log data, which may overlook subtle behavioral cues linked to learners' cognitive strategies. This paper proposes a multimodal late fusion model that integrates screencast-based visual data and structured in-game action sequences to classify students' problem-solving strategies. In a pilot study with secondary school students (N=149) playing a multitouch educational game, the fusion model outperformed unimodal baseline models, increasing classification accuracy by over 15%. Results highlight the potential of multimodal ML for strategy-sensitive assessment and adaptive support in interactive learning contexts.

📄 PDF Abstract BibTeX arXiv:2507.22426

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

HunyuanVideo-Foley: Multimodal Diffusion with Representation Alignment for High-Fidelity Foley Audio Generation

2025-08-23 · Sizhe Shan, Qiulin Li, Yutao Cui, Miles Yang 외 arxiv

Recent advances in video generation produce visually realistic content, yet the absence of synchronized audio severely compromises immersion. To address key challenges in video-to-audio generation, including multimodal d…

Audio GenerationVideo Generation

When Hillclimbers Beat Genetic Algorithms in Multimodal Optimization

2015-04-26 · Fernando G. Lobo, Mosab Bazargani

It has been shown in the past that a multistart hillclimbing strategy compares favourably to a standard genetic algorithm with respect to solving instances of the multimodal problem generator. We extend that work and ver…

Diversity

VAMPS: Visual-Assisted Mathematical Problem Solving Benchmark

2026-06-02 · Amirhossein Dabiriaghdam, Shayan Vassef, Mohammadreza Bakhtiari, Yasamin Medghalchi 외 arxiv

Multimodal large language models are increasingly capable of complex reasoning, yet their performance often degrades when they must externalize a problem through a tool and then reason over the tool's output, specificall…

An Attention-based Multi-Scale Feature Learning Network for Multimodal Medical Image Fusion

2022-12-09 · Meng Zhou, Xiaolan Xu, Yuxuan Zhang

Medical images play an important role in clinical applications. Multimodal medical images could provide rich information about patients for physicians to diagnose. The image fusion technique is able to synthesize complem…

Diagnostic

Hybrid Multimodal Fusion for Humor Detection

2022-09-24 · Haojie Xu, Weifeng Liu, Jingwei Liu, Mingzheng Li 외

In this paper, we present our solution to the MuSe-Humor sub-challenge of the Multimodal Emotional Challenge (MuSe) 2022. The goal of the MuSe-Humor sub-challenge is to detect humor and calculate AUC from audiovisual rec…

Humor Detection