paper-with-me

홈 › Papers

A model for interpreting social interactions in local image regions

2017-12-26 · Guy Ben-Yosef, Alon Yachin, Shimon Ullman

Understanding social interactions (such as 'hug' or 'fight') is a basic and important capacity of the human visual system, but a challenging and still open problem for modeling. In this work we study visual recognition of social interactions, based on small but recognizable local regions. The approach is based on two novel key components: (i) A given social interaction can be recognized reliably from reduced images (called 'minimal images'). (ii) The recognition of a social interaction depends on identifying components and relations within the minimal image (termed 'interpretation'). We show psychophysics data for minimal images and modeling results for their interpretation. We discuss the integration of minimal configurations in recognizing social interactions in a detailed, high-resolution image.

📄 PDF Abstract BibTeX arXiv:1712.09299

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

MINGLE: VLMs for Semantically Complex Region Detection in Urban Scenes

2025-09-16 · Liu Liu, Alexandra Kudaeva, Marco Cipriano, Fatimeh Al Ghannam 외 arxiv

Understanding group-level social interactions in public spaces is crucial for urban planning, informing the design of socially vibrant and inclusive environments. Detecting such interactions from images involves interpre…

Depth EstimationObject Detection

ISG: I can See Your Gene Expression

2022-10-30 · Yan Yang, Liyuan Pan, Liu Liu, Eric A Stone

This paper aims to predict gene expression from a histology slide image precisely. Such a slide image has a large resolution and sparsely distributed textures. These obstruct extracting and interpreting discriminative fe…

MimeQA: Towards Socially-Intelligent Nonverbal Foundation Models

2025-02-23 · Hengzhi Li, Megan Tjandrasuwita, Yi R. Fung, Armando Solar-Lezama 외

Socially intelligent AI that can understand and interact seamlessly with humans in daily lives is increasingly important as AI becomes more closely integrated with peoples' daily activities. However, current works in art…

Modeling Relationships in Referential Expressions with Compositional Modular Networks

2016-11-30 · CVPR 2017 7 · Ronghang Hu, Marcus Rohrbach, Jacob Andreas, Trevor Darrell 외

People often refer to entities in an image in terms of their relationships with other entities. For example, "the black cat sitting under the table" refers to both a "black cat" entity and its relationship with another "…

Visual Question Answering (VQA)

Multimodal Coreference Resolution for Chinese Social Media Dialogues: Dataset and Benchmark Approach

2025-04-19 · Xingyu Li, Chen Gong, Guohong Fu

Multimodal coreference resolution (MCR) aims to identify mentions referring to the same entity across different modalities, such as text and visuals, and is essential for understanding multimodal content. In the era of r…

coreference-resolutionCoreference Resolution