paper-with-me

홈 › Papers

Graph Neural Networks for Image Understanding Based on Multiple Cues: Group Emotion Recognition and Event Recognition as Use Cases

2019-09-19 · Xin Guo, Luisa F. Polania, Bin Zhu, Charles Boncelet, Kenneth E. Barner

A graph neural network (GNN) for image understanding based on multiple cues is proposed in this paper. Compared to traditional feature and decision fusion approaches that neglect the fact that features can interact and exchange information, the proposed GNN is able to pass information among features extracted from different models. Two image understanding tasks, namely group-level emotion recognition (GER) and event recognition, which are highly semantic and require the interaction of several deep models to synthesize multiple cues, were selected to validate the performance of the proposed method. It is shown through experiments that the proposed method achieves state-of-the-art performance on the selected image understanding tasks. In addition, a new group-level emotion recognition database is introduced and shared in this paper.

📄 PDF Abstract BibTeX arXiv:1909.12911

Code (1)

gxstudy/Graph-Neural-Networks-for-Image-Understanding-Based-on-Multiple-Cues 공식 구현 tf

Tasks

Emotion RecognitionGraph Neural Network

Methods 이 논문이 사용한 방법론

Graph Neural Network 설명 없음

Similar Papers 제목 키워드 기반

VIP: Finding Important People in Images

2015-02-19 · CVPR 2015 6 · Clint Solomon Mathialagan, Andrew C. Gallagher, Dhruv Batra

People preserve memories of events such as birthdays, weddings, or vacations by capturing photos, often depicting groups of people. Invariably, some individuals in the image are more important than others given the conte…

Different Demographic Cues Yield Inconsistent Conclusions About LLM Personalization and Bias

2026-01-26 · Manuel Tonneau, Neil K. R. Seghal, Niyati Malhotra, Sharif Kazemi 외 arxiv

Demographic cue-based evaluation is widely used to study how large language models (LLMs) adapt their responses to signaled demographic attributes within and across groups. This approach typically relies on a single cue …

Context-Aware Network Based on Multi-scale Spatio-temporal Attention for Action Recognition in Videos

2025-12-21 · Xiaoyang Li, Wenzhu Yang, Kanglin Wang, Tiebiao Wang 외 arxiv

Action recognition is a critical task in video understanding, requiring the comprehensive capture of spatio-temporal cues across various scales. However, existing methods often overlook the multi-granularity nature of ac…

Action Recognition In Videos

Face, Body, Voice: Video Person-Clustering with Multiple Modalities

2021-05-20 · Andrew Brown, Vicky Kalogeiton, Andrew Zisserman

The objective of this work is person-clustering in videos -- grouping characters according to their identity. Previous methods focus on the narrower task of face-clustering, and for the most part ignore other cues such a…

ClusteringFace Clustering

One Persona, Many Cues, Different Results: How Sociodemographic Cues Impact LLM Personalization

2026-01-26 · Franziska Weeber, Vera Neplenbroek, Jan Batzner, Sebastian Padó arxiv

Personalization of LLMs by sociodemographic subgroup often improves user experience, but can also introduce or amplify biases and unfair outcomes across groups. Prior work has employed so-called personas, sociodemographi…