paper-with-me

Audio Classification 벤치마크

Audio Classification on FSD50K

20개 결과 · ⬇ CSV · JSON

mAP

53.7 57.7 61.7 65.7 69.7 2021-02 2026-09 PSLA — 56.71 (2021-02-02) PSLA — 56.71 (2021-02-02) Large 6-Layer Transformer with Pooling — 53.7 (2021-05-01) Large 6-Layer Transformer with Pooling — 53.7 (2021-05-01) PaSST-S — 65.55 (2021-10-11) PaSST-N-S — 64.2 (2021-10-11) PaSST-S — 65.55 (2021-10-11) PaSST-N-S — 64.2 (2021-10-11) Temporal Knowledge Distillation for On-device Audio Classification — 54.8 (2021-10-27) Temporal Knowledge Distillation for On-device Audio Classification — 54.8 (2021-10-27) ONE-PEACE — 69.7 (2023-05-18) ONE-PEACE — 69.7 (2023-05-18) MN — 65.6 (2023-10-24) DyMN-L — 65.5 (2023-10-24) MN — 65.6 (2023-10-24) DyMN-L — 65.5 (2023-10-24) MATPAC (SSL Model) — 55.2 (2025-02-17) MATPAC (SSL Model) — 55.2 (2025-02-17) PSLA — 56.71 (2021-02-02) PaSST-S — 65.55 (2021-10-11) ONE-PEACE — 69.7 (2023-05-18)
RankModel mAPMean AP Extra Training Data PaperCodeYear
1 ONE-PEACE 69.7 ONE-PEACE: Exploring One General Representation Model Toward Unlimited Modalities modelscope/modelscope · OFA-Sys/ONE-PEACE 2023
2 MN 65.6 Dynamic Convolutional Neural Networks as Efficient Pre-trained Audio Models fschmid56/efficientat 2023
3 PaSST-S 65.55 Efficient Training of Audio Transformers with Patchout kkoutini/passt · kkoutini/passt_hear21 2021
4 DyMN-L 65.5 Dynamic Convolutional Neural Networks as Efficient Pre-trained Audio Models fschmid56/efficientat 2023
5 PaSST-N-S 64.2 Efficient Training of Audio Transformers with Patchout kkoutini/passt · kkoutini/passt_hear21 2021
6 PSLA 56.71 PSLA: Improving Audio Tagging with Pretraining, Sampling, Labeling, and Aggregation YuanGongND/psla 2021
7 MATPAC (SSL Model) 55.2 Masked Latent Prediction and Classification for Self-Supervised Audio Representation Learning aurianworld/matpac 2025
8 Temporal Knowledge Distillation for On-device Audio Classification 54.8 Temporal Knowledge Distillation for On-device Audio Classification 2021
9 Large 6-Layer Transformer with Pooling 53.7 Audio Transformers 2021
10 LHGNN 59 LHGNN: Local-Higher Order Graph Neural Networks For Audio Classification and Tagging 2025
11 ONE-PEACE 69.7 ONE-PEACE: Exploring One General Representation Model Toward Unlimited Modalities modelscope/modelscope · OFA-Sys/ONE-PEACE 2023
12 MN 65.6 Dynamic Convolutional Neural Networks as Efficient Pre-trained Audio Models fschmid56/efficientat 2023
13 PaSST-S 65.55 Efficient Training of Audio Transformers with Patchout kkoutini/passt · kkoutini/passt_hear21 2021
14 DyMN-L 65.5 Dynamic Convolutional Neural Networks as Efficient Pre-trained Audio Models fschmid56/efficientat 2023
15 PaSST-N-S 64.2 Efficient Training of Audio Transformers with Patchout kkoutini/passt · kkoutini/passt_hear21 2021
16 PSLA 56.71 PSLA: Improving Audio Tagging with Pretraining, Sampling, Labeling, and Aggregation YuanGongND/psla 2021
17 MATPAC (SSL Model) 55.2 Masked Latent Prediction and Classification for Self-Supervised Audio Representation Learning aurianworld/matpac 2025
18 Temporal Knowledge Distillation for On-device Audio Classification 54.8 Temporal Knowledge Distillation for On-device Audio Classification 2021
19 Large 6-Layer Transformer with Pooling 53.7 Audio Transformers 2021
20 LHGNN 59 LHGNN: Local-Higher Order Graph Neural Networks For Audio Classification and Tagging 2025
1–20 / 20 페이지당 10 20 50 100