paper-with-me

Natural Language Moment Retrieval 벤치마크

Natural Language Moment Retrieval on TACoS

13개 결과 · ⬇ CSV · JSON

R@1,IoU=0.3

45.46 48.62 51.78 54.94 58.1 2020-11 2026-09 VLG-Net — 45.46 (2020-11-19) GVL (paragraph-level) — 48.29 (2023-03-11) GVL — 45.92 (2023-03-11) UniVTG — 51.44 (2023-07-31) CG-DETR — 52.23 (2023-11-15) BAM-DETR — 56.69 (2023-11-30) LLMEPET — 52.73 (2024-07-21) SG-DETR (w/ PT) — 58.1 (2024-10-02) SG-DETR — 56.71 (2024-10-02) FlashVTG — 53.71 (2024-12-18) LD-DETR — 57.61 (2025-01-18) DeCafNet — 57.36 (2025-05-22) VLG-Net — 45.46 (2020-11-19) GVL (paragraph-level) — 48.29 (2023-03-11) UniVTG — 51.44 (2023-07-31) CG-DETR — 52.23 (2023-11-15) BAM-DETR — 56.69 (2023-11-30) SG-DETR (w/ PT) — 58.1 (2024-10-02)
RankModel R@1,IoU=0.3R@1,IoU=0.5R@1,IoU=0.7R@5,IoU=0.1R@5,IoU=0.3R@5,IoU=0.5mIoU Extra Training Data PaperCodeYear
1 SG-DETR (w/ PT) 58.1046.4033.9042.40 Saliency-Guided DETR for Moment Retrieval and Highlight Detection ai-forever/sg-detr 2024
2 LD-DETR 57.61 44.3126.2440.30 LD-DETR: Loop Decoder DEtection TRansformer for Video Moment Retrieval and Highlight Detection qingchen239/ld-detr 2025
3 DeCafNet 57.3646.7981.0571.13 DeCafNet: Delegate and Conquer for Efficient Temporal Grounding in Long Videos zijialewislu/cvpr2025-decafnet 2025
4 SG-DETR 56.7144.7029.9040.90 Saliency-Guided DETR for Moment Retrieval and Highlight Detection ai-forever/sg-detr 2024
5 BAM-DETR 56.6941.5426.7739.31 BAM-DETR: Boundary-Aligned Moment Detection Transformer for Temporal Sentence Grounding in Videos Pilhyeon/BAM-DETR 2023
6 FlashVTG 53.7141.7624.7437.61 FlashVTG: Feature Layering and Adaptive Score Handling Network for Video Temporal Grounding zhuo-cao/flashvtg 2024
7 LLMEPET 52.7340.1222.7836.55 Prior Knowledge Integration via LLM Encoding and Pseudo Event Regulation for Video Moment Retrieval fletcherjiang/llmepet 2024
8 CG-DETR 52.2339.6122.2336.48 Correlation-Guided Query-Dependency Calibration for Video Temporal Grounding wjun0830/qd-detr · wjun0830/cgdetr 2023
9 UniVTG 51.4434.9721.0735.76 UniVTG: Towards Unified Video-Language Temporal Grounding showlab/univtg 2023
10 GVL (paragraph-level) 48.2936.07 Learning Grounded Vision-Language Representation for Versatile Understanding in Untrimmed Videos zjr2000/gvl 2023
11 GVL 45.9234.57 Learning Grounded Vision-Language Representation for Versatile Understanding in Untrimmed Videos zjr2000/gvl 2023
12 VLG-Net 45.4634.1981.8070.3856.56 VLG-Net: Video-Language Graph Matching Network for Video Grounding Soldelli/VLG-Net 2020
13 UVCOM 36.3923.32 Bridging the Gap: A Unified Video Comprehension Framework for Moment Retrieval and Highlight Detection easonxiao-888/uvcom 2023
1–13 / 13 페이지당 10 20 50 100