paper-with-me

Sound Prompted Semantic Segmentation 벤치마크

Sound Prompted Semantic Segmentation on ADE20K

8개 결과 · ⬇ CSV · JSON

mAP

16.8 20.78 24.75 28.73 32.7 2018-04 2026-09 DAVENet — 16.8 (2018-04-04) DAVENet — 16.8 (2018-04-04) CAVMAE — 26.0 (2022-10-02) CAVMAE — 26.0 (2022-10-02) ImageBIND — 19.7 (2023-05-09) ImageBIND — 19.7 (2023-05-09) DenseAV — 32.7 (2024-06-09) DenseAV — 32.7 (2024-06-09) DAVENet — 16.8 (2018-04-04) CAVMAE — 26.0 (2022-10-02) DenseAV — 32.7 (2024-06-09)
RankModel mAPmIoU PaperCodeYear
1 DenseAV 32.724.7 Separating the "Chirp" from the "Chat": Self-supervised Visual Grounding of Sound and Language mhamilton723/DenseAV 2024
2 CAVMAE 26.017.0 Contrastive Audio-Visual Masked Autoencoder yuangongnd/cav-mae 2022
3 ImageBIND 19.720.5 ImageBind: One Embedding Space To Bind Them All facebookresearch/imagebind · klemens-floege/oneprot · ginihumer/amumo 2023
4 DAVENet 16.818.1 Jointly Discovering Visual Objects and Spoken Words from Raw Sensory Input 2018
5 DenseAV 32.724.7 Separating the "Chirp" from the "Chat": Self-supervised Visual Grounding of Sound and Language mhamilton723/DenseAV 2024
6 CAVMAE 26.017.0 Contrastive Audio-Visual Masked Autoencoder yuangongnd/cav-mae 2022
7 ImageBIND 19.720.5 ImageBind: One Embedding Space To Bind Them All facebookresearch/imagebind · klemens-floege/oneprot · ginihumer/amumo 2023
8 DAVENet 16.818.1 Jointly Discovering Visual Objects and Spoken Words from Raw Sensory Input 2018
1–8 / 8 페이지당 10 20 50 100