paper-with-me

Zero-Shot Transfer Image Classification 벤치마크

Zero-Shot Transfer Image Classification on ImageNet-A

13개 결과 · ⬇ CSV · JSON

Accuracy (Private)

44.7 56.08 67.45 78.83 90.2 2021-02 2026-09 ALIGN — 75.8 (2021-02-11) CLIP — 77.2 (2021-02-26) LiT-tuning — 79.4 (2021-11-15) BASIC — 85.6 (2021-11-19) CoCa — 90.2 (2022-05-04) LiT ViT-e — 88.0 (2022-09-14) PaLI — 44.7 (2022-09-14) AltCLIP — 69.5 (2022-11-12) LiT-22B — 90.1 (2023-02-10) EVA-CLIP-E/14+ — 82.1 (2023-03-27) InternVL-C — 83.8 (2023-12-21) EVA-CLIP-18B — 87.3 (2024-02-06) ALIGN — 75.8 (2021-02-11) CLIP — 77.2 (2021-02-26) LiT-tuning — 79.4 (2021-11-15) BASIC — 85.6 (2021-11-19) CoCa — 90.2 (2022-05-04)
RankModel Accuracy (Private)Accuracy (Public) PaperCodeYear
1 CoCa 90.2 CoCa: Contrastive Captioners are Image-Text Foundation Models mlfoundations/open_clip · facebookresearch/multimodal · lucidrains/CoCa-pytorch · +3 2022
2 LiT-22B 90.1 Scaling Vision Transformers to 22 Billion Parameters lucidrains/flash-cosine-sim-attention 2023
3 LiT ViT-e 88.0 PaLI: A Jointly-Scaled Multilingual Language-Image Model google-research/big_vision 2022
4 EVA-CLIP-18B 87.3 EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters baaivision/EVA · baaivision/eva 2024
5 BASIC (Lion) 86.4
6 BASIC 85.6 Combined Scaling for Zero-shot Transfer Learning 2021
7 InternVL-C 83.8 InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks opengvlab/internvl · opengvlab/internvl-mmdetseg 2023
8 EVA-CLIP-E/14+ 82.1 EVA-CLIP: Improved Training Techniques for CLIP at Scale baaivision/eva · PaddlePaddle/PaddleMIX · Yui010206/CREMA · +1 2023
9 LiT-tuning 79.4 37.8 LiT: Zero-Shot Transfer with Locked-image text Tuning mlfoundations/open_clip · google-research/vision_transformer · google-research/big_vision · +2 2021
10 CLIP 77.2- Learning Transferable Visual Models From Natural Language Supervision openai/CLIP · mlfoundations/open_clip · towhee-io/towhee · +79 2021
11 ALIGN 75.8- Scaling Up Visual and Vision-Language Representation Learning With Noisy Text Supervision facebookresearch/metaclip · kakaobrain/coyo-dataset · MicPie/clasp · +2 2021
12 AltCLIP 69.5 AltCLIP: Altering the Language Encoder in CLIP for Extended Language Capabilities flagai-open/flagai · pwc-1/Paper-8 2022
13 PaLI 44.7 PaLI: A Jointly-Scaled Multilingual Language-Image Model google-research/big_vision 2022
1–13 / 13 페이지당 10 20 50 100