paper-with-me

홈 › Papers

Multimodal Methods for Analyzing Learning and Training Environments: A Systematic Literature Review

2024-08-22 · Clayton Cohn, Eduardo Davalos, Caleb Vatral, Joyce Horn Fonteles, Hanchen David Wang, Meiyi Ma, Gautam Biswas

Recent technological advancements have enhanced our ability to collect and analyze rich multimodal data (e.g., speech, video, and eye gaze) to better inform learning and training experiences. While previous reviews have focused on parts of the multimodal pipeline (e.g., conceptual models and data fusion), a comprehensive literature review on the methods informing multimodal learning and training environments has not been conducted. This literature review provides an in-depth analysis of research methods in these environments, proposing a taxonomy and framework that encapsulates recent methodological advances in this field and characterizes the multimodal domain in terms of five modality groups: Natural Language, Video, Sensors, Human-Centered, and Environment Logs. We introduce a novel data fusion category -- mid fusion -- and a graph-based technique for refining literature reviews, termed citation graph pruning. Our analysis reveals that leveraging multiple modalities offers a more holistic understanding of the behaviors and outcomes of learners and trainees. Even when multimodality does not enhance predictive accuracy, it often uncovers patterns that contextualize and elucidate unimodal data, revealing subtleties that a single modality may miss. However, there remains a need for further research to bridge the divide between multimodal learning and training studies and foundational AI research.

📄 PDF Abstract BibTeX arXiv:2408.14491

Code (0)

등록된 구현이 없습니다.

Tasks

Systematic Literature Review

Similar Papers 제목 키워드 기반

LLMGeo: Benchmarking Large Language Models on Image Geolocation In-the-wild

2024-05-30 · Zhiqiang Wang, Dejia Xu, Rana Muhammad Shahroz Khan, Yanbin Lin 외

Image geolocation is a critical task in various image-understanding applications. However, existing methods often fail when analyzing challenging, in-the-wild images. Inspired by the exceptional background knowledge of m…

Benchmarking

A First Step in Using Machine Learning Methods to Enhance Interaction Analysis for Embodied Learning Environments

2024-05-10 · Joyce Fonteles, Eduardo Davalos, Ashwin T. S., Yike Zhang 외

Investigating children's embodied learning in mixed-reality environments, where they collaboratively simulate scientific processes, requires analyzing complex multimodal data to interpret their learning and coordination …

Mixed Reality

Adaptation through prediction: multisensory active inference torque control

2021-12-13 · Cristian Meo, Giovanni Franzese, Corrado Pezzato, Max Spahn 외

Adaptation to external and internal changes is major for robotic systems in uncertain environments. Here we present a novel multisensory active inference torque controller for industrial arms that shows how prediction ca…

Prediction

Online Continual Learning: A Systematic Literature Review of Approaches, Challenges, and Benchmarks

2025-01-09 · Seyed Amir Bidaki, Amir Mohammadkhah, Kiyan Rezaee, Faeze Hassani 외

Online Continual Learning (OCL) is a critical area in machine learning, focusing on enabling models to adapt to evolving data streams in real-time while addressing challenges such as catastrophic forgetting and the stabi…

Continual Learningimage-classificationImage Classificationobject-detection+3

UniFinEval: Towards Unified Evaluation of Financial Multimodal Models across Text, Images and Videos

2026-01-09 · Zhi Yang, Lingfeng Zeng, Fangqi Lou, Qi Qi 외 arxiv

Multimodal large language models are playing an increasingly significant role in empowering the financial domain, however, the challenges they face, such as multimodal and high-density information and cross-modal multi-h…