paper-with-me

Highlight Detection

4개 벤치마크 · 논문 93편 · 이 태스크의 논문 보기 →

Benchmarks

QVHighlights

결과 43개

TvSum

결과 14개

YouTube Highlights

결과 14개

arabiska

결과 2개

Most implemented

Papers

SVHighlights: Towards Extremely Long Sport Video Highlight Detection

2026-06-05 · Donggyu Lee, Youngbin Ki, Jeonghun Kang, Taehwan Kim arxiv

While highlight detection for long-form videos is of great practical importance, most existing methods remain limited to short-form content, largely due to the absence of a suitable benchmark. To bridge this gap, we intr…

Highlight Detection

Turing Patterns for Multimedia: Reaction-Diffusion Multi-Modal Fusion for Language-Guided Video Moment Retrieval

2026-06-01 · Xiang Fang, Wanlong Fang, Wei Ji, Tat-Seng Chua arxiv

Video-language models are pivotal for tasks such as moment retrieval and highlight detection, yet they often struggle to capture the dynamic, non-linear interactions between temporal video sequences and textual semantics…

Saliency PredictionHighlight DetectionMoment Retrieval

CoSTL: Comprehensive Spatial-Temporal Representation Learning for Moment Retrieval and Highlight Detection

2026-05-31 · Xin Dong, Wenjia Geng, Wenfeng Deng, Yansong Tang arxiv

Video Moment Retrieval (MR) and Highlight Detection (HD) are crucial tasks in video analysis that aim to localize specific moments and estimate clip-wise relevance based on a given text query. Recent approaches treat the…

Representation LearningHighlight DetectionMoment RetrievalVideo Grounding

GroundVTS: Visual Token Sampling in Multimodal Large Language Models for Video Temporal Grounding

2026-04-02 · Rong Fan, Kaiyan Xiao, Minghao Zhu, Liuyi Wang 외 arxiv

Video temporal grounding (VTG) is a critical task in video understanding and a key capability for extending video large language models (Vid-LLMs) to broader applications. However, existing Vid-LLMs rely on uniform frame…

Highlight DetectionMoment Retrieval

Follow the Saliency: Supervised Saliency for Retrieval-augmented Dense Video Captioning

2026-03-12 · Seung hee Choi, MinJu Jeon, Hyunwoo Oh, Jihwan Lee 외 arxiv

Existing retrieval-augmented approaches for Dense Video Captioning (DVC) often fail to achieve accurate temporal segmentation aligned with true event boundaries, as they rely on heuristic strategies that overlook ground …

Dense Video CaptioningHighlight Detection

Sounding Highlights: Dual-Pathway Audio Encoders for Audio-Visual Video Highlight Detection

2026-02-03 · Seohyun Joo, Yoori Oh arxiv

Audio-visual video highlight detection aims to automatically identify the most salient moments in videos by leveraging both visual and auditory cues. However, existing models often underutilize the audio modality, focusi…

Highlight Detection

전체 93편 보기 →