paper-with-me

홈 › Papers

EnTri: Ensemble Learning with Tri-level Representations for Explainable Scene Recognition

2023-07-23 · Amirhossein Aminimehr, Amirali Molaei, Erik Cambria

Scene recognition based on deep-learning has made significant progress, but there are still limitations in its performance due to challenges posed by inter-class similarities and intra-class dissimilarities. Furthermore, prior research has primarily focused on improving classification accuracy, yet it has given less attention to achieving interpretable, precise scene classification. Therefore, we are motivated to propose EnTri, an ensemble scene recognition framework that employs ensemble learning using a hierarchy of visual features. EnTri represents features at three distinct levels of detail: pixel-level, semantic segmentation-level, and object class and frequency level. By incorporating distinct feature encoding schemes of differing complexity and leveraging ensemble strategies, our approach aims to improve classification accuracy while enhancing transparency and interpretability via visual and textual explanations. To achieve interpretability, we devised an extension algorithm that generates both visual and textual explanations highlighting various properties of a given scene that contribute to the final prediction of its category. This includes information about objects, statistics, spatial layout, and textural details. Through experiments on benchmark scene classification datasets, EnTri has demonstrated superiority in terms of recognition accuracy, achieving competitive performance compared to state-of-the-art approaches, with an accuracy of 87.69%, 75.56%, and 99.17% on the MIT67, SUN397, and UIUC8 datasets, respectively.

📄 PDF Abstract BibTeX arXiv:2307.12442

Code (0)

등록된 구현이 없습니다.

Tasks

ClassificationEnsemble LearningScene ClassificationScene RecognitionSemantic Segmentation

Similar Papers 제목 키워드 기반

Scene-Agnostic Object-Centric Representation Learning for 3D Gaussian Splatting

2026-04-10 · Tsuheng Hsu, Guiyu Liu, Juho Kannala, Janne Heikkilä arxiv

Recent works on 3D scene understanding leverage 2D masks from visual foundation models (VFMs) to supervise radiance fields, enabling instance-level 3D segmentation. However, the supervision signals from foundation models…

Representation LearningScene Understanding

EVAL: Explainable Video Anomaly Localization

2022-12-15 · CVPR 2023 1 · Ashish Singh, Michael J. Jones, Erik Learned-Miller

We develop a novel framework for single-scene video anomaly localization that allows for human-understandable reasons for the decisions the system makes. We first learn general representations of objects and their motion…

Anomaly DetectionAnomaly LocalizationVideo Anomaly Detection

Learning Global Spatial Information for Multi-View Object-Centric Models

2021-09-29 · Yuya Kobayashi, Masahiro Suzuki, Yutaka Matsuo

Recently, several studies have been working on multi-view object-centric models, which predict unobserved views of a scene and infer object-centric representations from several observation views. In general, multi-object…

Novel View SynthesisObject

Slot-VAE: Object-Centric Scene Generation with Slot Attention

2023-06-12 · Yanbo Wang, Letao Liu, Justin Dauwels

Slot attention has shown remarkable object-centric representation learning performance in computer vision tasks without requiring any supervision. Despite its object-centric binding ability brought by compositional model…

ObjectRepresentation LearningScene Generation

CARFF: Conditional Auto-encoded Radiance Field for 3D Scene Forecasting

2024-01-31 · Jiezhi Yang, Khushi Desai, Charles Packer, Harshil Bhatia 외

We propose CARFF, a method for predicting future 3D scenes given past observations. Our method maps 2D ego-centric images to a distribution over plausible 3D latent scene configurations and predicts the evolution of hypo…

Autonomous DrivingNeRFNeural Rendering