paper
-with-
me
Papers
Browse State-of-the-Art
Datasets
Methods
AI Agents
Trends
Digest
🌙
Audio Classification
벤치마크
Audio Classification on
FSD50K
20개 결과 ·
⬇ CSV
·
JSON
mAP
53.7
57.7
61.7
65.7
69.7
2021-02
2026-09
PSLA — 56.71 (2021-02-02)
PSLA — 56.71 (2021-02-02)
Large 6-Layer Transformer with Pooling — 53.7 (2021-05-01)
Large 6-Layer Transformer with Pooling — 53.7 (2021-05-01)
PaSST-S — 65.55 (2021-10-11)
PaSST-N-S — 64.2 (2021-10-11)
PaSST-S — 65.55 (2021-10-11)
PaSST-N-S — 64.2 (2021-10-11)
Temporal Knowledge Distillation for On-device Audio Classification — 54.8 (2021-10-27)
Temporal Knowledge Distillation for On-device Audio Classification — 54.8 (2021-10-27)
ONE-PEACE — 69.7 (2023-05-18)
ONE-PEACE — 69.7 (2023-05-18)
MN — 65.6 (2023-10-24)
DyMN-L — 65.5 (2023-10-24)
MN — 65.6 (2023-10-24)
DyMN-L — 65.5 (2023-10-24)
MATPAC (SSL Model) — 55.2 (2025-02-17)
MATPAC (SSL Model) — 55.2 (2025-02-17)
PSLA — 56.71 (2021-02-02)
PaSST-S — 65.55 (2021-10-11)
ONE-PEACE — 69.7 (2023-05-18)
2021-02-02 — PSLA: mAP 56.71
2021-10-11 — PaSST-S: mAP 65.55
2023-05-18 — ONE-PEACE: mAP 69.7
Rank
Model
mAP
Mean AP
Extra Training Data
Paper
Code
Year
1
ONE-PEACE
69.7
–
✓
ONE-PEACE: Exploring One General Representation Model Toward Unlimited Modalities
modelscope/modelscope
·
OFA-Sys/ONE-PEACE
2023
2
MN
65.6
–
✓
Dynamic Convolutional Neural Networks as Efficient Pre-trained Audio Models
fschmid56/efficientat
2023
3
PaSST-S
65.55
–
✓
Efficient Training of Audio Transformers with Patchout
kkoutini/passt
·
kkoutini/passt_hear21
2021
4
DyMN-L
65.5
–
✓
Dynamic Convolutional Neural Networks as Efficient Pre-trained Audio Models
fschmid56/efficientat
2023
5
PaSST-N-S
64.2
–
✓
Efficient Training of Audio Transformers with Patchout
kkoutini/passt
·
kkoutini/passt_hear21
2021
6
PSLA
56.71
–
✓
PSLA: Improving Audio Tagging with Pretraining, Sampling, Labeling, and Aggregation
YuanGongND/psla
2021
7
MATPAC (SSL Model)
55.2
–
Masked Latent Prediction and Classification for Self-Supervised Audio Representation Learning
aurianworld/matpac
2025
8
Temporal Knowledge Distillation for On-device Audio Classification
54.8
–
Temporal Knowledge Distillation for On-device Audio Classification
2021
9
Large 6-Layer Transformer with Pooling
53.7
–
Audio Transformers
2021
10
LHGNN
–
59
LHGNN: Local-Higher Order Graph Neural Networks For Audio Classification and Tagging
2025
11
ONE-PEACE
69.7
–
✓
ONE-PEACE: Exploring One General Representation Model Toward Unlimited Modalities
modelscope/modelscope
·
OFA-Sys/ONE-PEACE
2023
12
MN
65.6
–
✓
Dynamic Convolutional Neural Networks as Efficient Pre-trained Audio Models
fschmid56/efficientat
2023
13
PaSST-S
65.55
–
✓
Efficient Training of Audio Transformers with Patchout
kkoutini/passt
·
kkoutini/passt_hear21
2021
14
DyMN-L
65.5
–
✓
Dynamic Convolutional Neural Networks as Efficient Pre-trained Audio Models
fschmid56/efficientat
2023
15
PaSST-N-S
64.2
–
✓
Efficient Training of Audio Transformers with Patchout
kkoutini/passt
·
kkoutini/passt_hear21
2021
16
PSLA
56.71
–
✓
PSLA: Improving Audio Tagging with Pretraining, Sampling, Labeling, and Aggregation
YuanGongND/psla
2021
17
MATPAC (SSL Model)
55.2
–
Masked Latent Prediction and Classification for Self-Supervised Audio Representation Learning
aurianworld/matpac
2025
18
Temporal Knowledge Distillation for On-device Audio Classification
54.8
–
Temporal Knowledge Distillation for On-device Audio Classification
2021
19
Large 6-Layer Transformer with Pooling
53.7
–
Audio Transformers
2021
20
LHGNN
–
59
LHGNN: Local-Higher Order Graph Neural Networks For Audio Classification and Tagging
2025
1–20 / 20
페이지당
10
20
50
100