paper-with-me

GZSL Video Classification

6개 벤치마크 · 논문 7편 · 이 태스크의 논문 보기 →

Benchmarks

UCF-GZSL(main)

결과 7개

VGGSound-GZSL(main)

결과 7개

UCF-GZSL (cls)

결과 4개

VGGSound-GZSL (cls)

결과 4개

Most implemented

Papers

Boosting Audio-visual Zero-shot Learning with Large Language Models

2023-11-21 · Haoxing Chen, Yaohui Li, Yan Hong, Zizheng Huang 외

Audio-visual zero-shot learning aims to recognize unseen classes based on paired audio-visual sequences. Recent methods mainly focus on learning multi-modal features aligned with class names to enhance the generalization…

audio-visual learningDescriptiveGZSL Video ClassificationZero-Shot Learning

Hyperbolic Audio-visual Zero-shot Learning

2023-08-24 · ICCV 2023 1 · Jie Hong, Zeeshan Hayder, Junlin Han, Pengfei Fang 외

Audio-visual zero-shot learning aims to classify samples consisting of a pair of corresponding audio and video sequences from classes that are not present during training. An analysis of the audio-visual data reveals a l…

GZSL Video ClassificationZero-Shot Learning

Temporal and cross-modal attention for audio-visual zero-shot learning

2022-07-20 · Otniel-Bogdan Mercea, Thomas Hummel, A. Sophia Koepke, Zeynep Akata

Audio-visual generalised zero-shot learning for video classification requires understanding the relations between the audio and visual information in order to be able to recognise samples from novel, previously unseen cl…

GZSL Video ClassificationVideo ClassificationZero-Shot Learning

Attribute Prototype Network for Any-Shot Learning

2022-04-04 · Wenjia Xu, Yongqin Xian, Jiuniu Wang, Bernt Schiele 외

Any-shot image classification allows to recognize novel classes with only a few or even zero samples. For the task of zero-shot learning, visual attributes have been shown to play an important role, while in the few-shot…

AttributeFew-Shot Image ClassificationGZSL Video Classificationimage-classification+3

Audio-visual Generalised Zero-shot Learning with Cross-modal Attention and Language

2022-03-07 · CVPR 2022 1 · Otniel-Bogdan Mercea, Lukas Riesch, A. Sophia Koepke, Zeynep Akata

Learning to classify video data from classes not included in the training data, i.e. video-based zero-shot learning, is challenging. We conjecture that the natural alignment between the audio and visual modalities in vid…

GZSL Video ClassificationZero-Shot LearningZSL Video Classification

AVGZSLNet: Audio-Visual Generalized Zero-Shot Learning by Reconstructing Label Features from Multi-Modal Embeddings

2020-05-27 · Pratik Mazumder, Pravendra Singh, Kranti Kumar Parida, Vinay P. Namboodiri

In this paper, we propose a novel approach for generalized zero-shot learning in a multi-modal setting, where we have novel classes of audio/video during testing that are not seen during training. We use the semantic rel…

DecoderGeneralized Zero-Shot LearningGZSL Video ClassificationRetrieval+3

전체 7편 보기 →