paper-with-me

홈 › Papers

Learning Hierarchical Visual Representations in Deep Neural Networks Using Hierarchical Linguistic Labels

2018-05-19 · Joshua C. Peterson, Paul Soulos, Aida Nematzadeh, Thomas L. Griffiths

Modern convolutional neural networks (CNNs) are able to achieve human-level object classification accuracy on specific tasks, and currently outperform competing models in explaining complex human visual representations. However, the categorization problem is posed differently for these networks than for humans: the accuracy of these networks is evaluated by their ability to identify single labels assigned to each image. These labels often cut arbitrarily across natural psychological taxonomies (e.g., dogs are separated into breeds, but never jointly categorized as "dogs"), and bias the resulting representations. By contrast, it is common for children to hear both "dog" and "Dalmatian" to describe the same stimulus, helping to group perceptually disparate objects (e.g., breeds) into a common mental class. In this work, we train CNN classifiers with multiple labels for each image that correspond to different levels of abstraction, and use this framework to reproduce classic patterns that appear in human generalization behavior.

📄 PDF Abstract BibTeX arXiv:1805.07647

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Hyperbolic Interaction Model For Hierarchical Multi-Label Classification

2019-05-26 · Boli Chen, Xin Huang, Lin Xiao, Zixin Cai 외

Different from the traditional classification tasks which assume mutual exclusion of labels, hierarchical multi-label classification (HMLC) aims to assign multiple labels to every instance with the labels organized under…

ClassificationGeneral ClassificationHierarchical Multi-label Classificationmodel+2

Learning Point-Language Hierarchical Alignment for 3D Visual Grounding

2022-10-22 · Jiaming Chen, Weixin Luo, Ran Song, Xiaolin Wei 외

This paper presents a novel hierarchical alignment model (HAM) that learns multi-granularity visual and linguistic representations in an end-to-end manner. We extract key points and proposal points to model 3D contexts a…

3D visual groundingSentenceVisual GroundingVocal Bursts Intensity Prediction

Learning Visual Hierarchies with Hyperbolic Embeddings

2024-11-26 · Ziwei Wang, Sameera Ramasinghe, Chenchen Xu, Julien Monteil 외

Structuring latent representations in a hierarchical manner enables models to learn patterns at multiple levels of abstraction. However, most prevalent image understanding models focus on visual similarity, and learning …

Image RetrievalRetrieval

HSVLT: Hierarchical Scale-Aware Vision-Language Transformer for Multi-Label Image Classification

2024-07-23 · Shuyi Ouyang, Hongyi Wang, Ziwei Niu, Zhenjia Bai 외

The task of multi-label image classification involves recognizing multiple objects within a single image. Considering both valuable semantic information contained in the labels and essential visual features presented in …

image-classificationImage ClassificationMulti-Label Image Classification

Hierarchical Lexical Manifold Projection in Large Language Models: A Novel Mechanism for Multi-Scale Semantic Representation

2025-02-08 · Natasha Martus, Sebastian Crowther, Maxwell Dorrington, Jonathan Applethwaite 외

The integration of structured hierarchical embeddings into transformer-based architectures introduces a refined approach to lexical representation, ensuring that multi-scale semantic relationships are preserved without c…

Adversarial TextComputational Efficiency