paper-with-me

Papers

Visual Superordinate Abstraction for Robust Concept Learning

2022-05-28 · Qi Zheng, Chaoyue Wang, Dadong Wang, DaCheng Tao

Concept learning constructs visual representations that are connected to linguistic semantics, which is fundamental to vision-language tasks. Although promising progress has been made, existing concept learners are still vulnerable to attribute perturbations and out-of-distribution compositions during inference. We ascribe the bottleneck to a failure of exploring the intrinsic semantic hierarchy of visual concepts, e.g. \{red, blue,...\} $\in$ color' subspace yet cube $\in$ shape'. In this paper, we propose a visual superordinate abstraction framework for explicitly modeling semantic-aware visual subspaces (i.e. visual superordinates). With only natural visual question answering data, our model first acquires the semantic hierarchy from a linguistic view, and then explores mutually exclusive visual superordinates under the guidance of linguistic hierarchy. In addition, a quasi-center visual concept clustering and a superordinate shortcut learning schemes are proposed to enhance the discrimination and independence of concepts within each visual superordinate. Experiments demonstrate the superiority of the proposed framework under diverse settings, which increases the overall answering accuracy relatively by 7.5\% on reasoning with perturbations and 15.6\% on compositional generalization tests.

📄 PDF Abstract BibTeX arXiv:2205.14444

Code (0)

등록된 구현이 없습니다.

Tasks

AttributeQuestion AnsweringVisual Question AnsweringVisual Question Answering (VQA)

Similar Papers 제목 키워드 기반

Object categorization in finer levels requires higher spatial frequencies, and therefore takes longer

2017-03-29 · Matin N. Ashtiani, Saeed Reza Kheradpisheh, Timothée Masquelier, Mohammad Ganjtabesh

The human visual system contains a hierarchical sequence of modules that take part in visual perception at different levels of abstraction, i.e., superordinate, basic, and subordinate levels. One important question is to…

Object CategorizationObject Recognition

From Concrete to Abstract: A Multimodal Generative Approach to Abstract Concept Learning

2024-10-03 · Haodong Xie, Rahul Singh Maharjan, Federico Tavella, Angelo Cangelosi

Understanding and manipulating concrete and abstract concepts is fundamental to human intelligence. Yet, they remain challenging for artificial agents. This paper introduces a multimodal generative approach to high order…

Investigating Concept Alignment Using Implausible Category Members

2026-05-20 · Sunayana Rane, Brenden M. Lake, Thomas L. Griffiths arxiv

Developing AI systems with a human-like understanding of everyday concepts is a key step towards developing safe, reliable systems whose behavior makes sense to humans. When probing concept understanding, asking question…

Effects of Linguistic Labels on Learned Visual Representations in Convolutional Neural Networks: Labels matter!

2019-09-25 · Seoyoung Ahn, Gregory Zelinsky, Gary Lupyan

We investigated the changes in visual representations learnt by CNNs when using different linguistic labels (e.g., trained with basic-level labels only, superordinate-level only, or both at the same time) and how they co…

Odd One Out

Key-Locked Rank One Editing for Text-to-Image Personalization

2023-05-02 · Yoad Tewel, Rinon Gal, Gal Chechik, Yuval Atzmon

Text-to-image models (T2I) offer a new level of flexibility by allowing users to guide the creative process through natural language. However, personalizing these models to align with user-provided visual concepts remain…