paper-with-me

Papers

Interpretable Concept-based Deep Learning Framework for Multimodal Human Behavior Modeling

2025-02-14 · Xinyu Li, Marwa Mahmoud

In the contemporary era of intelligent connectivity, Affective Computing (AC), which enables systems to recognize, interpret, and respond to human behavior states, has become an integrated part of many AI systems. As one of the most critical components of responsible AI and trustworthiness in all human-centered systems, explainability has been a major concern in AC. Particularly, the recently released EU General Data Protection Regulation requires any high-risk AI systems to be sufficiently interpretable, including biometric-based systems and emotion recognition systems widely used in the affective computing field. Existing explainable methods often compromise between interpretability and performance. Most of them focus only on highlighting key network parameters without offering meaningful, domain-specific explanations to the stakeholders. Additionally, they also face challenges in effectively co-learning and explaining insights from multimodal data sources. To address these limitations, we propose a novel and generalizable framework, namely the Attention-Guided Concept Model (AGCM), which provides learnable conceptual explanations by identifying what concepts that lead to the predictions and where they are observed. AGCM is extendable to any spatial and temporal signals through multimodal concept alignment and co-learning, empowering stakeholders with deeper insights into the model's decision-making process. We validate the efficiency of AGCM on well-established Facial Expression Recognition benchmark datasets while also demonstrating its generalizability on more complex real-world human behavior understanding applications.

📄 PDF Abstract BibTeX arXiv:2502.10145

Code (0)

등록된 구현이 없습니다.

Tasks

Concept AlignmentEmotion RecognitionFacial Expression Recognition

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

Probabilistic Concept Graph Reasoning for Multimodal Misinformation Detection

2026-03-26 · Ruichao Yang, Wei Gao, Xiaobin Zhu, Jing Ma 외 arxiv

Multimodal misinformation poses an escalating challenge that often evades traditional detectors, which are opaque black boxes and fragile against new manipulation tactics. We present Probabilistic Concept Graph Reasoning…

Human-like object concept representations emerge naturally in multimodal large language models

2024-07-01 · Changde Du, Kaicheng Fu, Bincheng Wen, Yi Sun 외

Understanding how humans conceptualize and categorize natural objects offers critical insights into perception and cognition. With the advent of Large Language Models (LLMs), a key question arises: can these models devel…

Triplet

Variational Information Pursuit with Large Language and Multimodal Models for Interpretable Predictions

2023-08-24 · Kwan Ho Ryan Chan, Aditya Chattopadhyay, Benjamin David Haeffele, Rene Vidal

Variational Information Pursuit (V-IP) is a framework for making interpretable predictions by design by sequentially selecting a short chain of task-relevant, user-defined and interpretable queries about the data that ar…

Semantic SimilaritySemantic Textual Similarity

Towards Faithful Multimodal Concept Bottleneck Models

2026-03-13 · Pierre Moreau, Emeline Pineau Ferrand, Yann Choho, Benjamin Wong 외 arxiv

Concept Bottleneck Models (CBMs) are interpretable models that route predictions through a layer of human-interpretable concepts. While widely studied in vision and, more recently, in NLP, CBMs remain largely unexplored …

Toward Human-AI Alignment in Large-Scale Multi-Player Games

2024-02-05 · Sugandha Sharma, Guy Davidson, Khimya Khetarpal, Anssi Kanervisto 외

Achieving human-AI alignment in complex multi-agent games is crucial for creating trustworthy AI agents that enhance gameplay. We propose a method to evaluate this alignment using an interpretable task-sets framework, fo…

AI Agent