paper-with-me

홈 › Papers

T-DEED: Temporal-Discriminability Enhancer Encoder-Decoder for Precise Event Spotting in Sports Videos

2024-04-08 · Artur Xarles, Sergio Escalera, Thomas B. Moeslund, Albert Clapés

In this paper, we introduce T-DEED, a Temporal-Discriminability Enhancer Encoder-Decoder for Precise Event Spotting in sports videos. T-DEED addresses multiple challenges in the task, including the need for discriminability among frame representations, high output temporal resolution to maintain prediction precision, and the necessity to capture information at different temporal scales to handle events with varying dynamics. It tackles these challenges through its specifically designed architecture, featuring an encoder-decoder for leveraging multiple temporal scales and achieving high output temporal resolution, along with temporal modules designed to increase token discriminability. Leveraging these characteristics, T-DEED achieves SOTA performance on the FigureSkating and FineDiving datasets. Code is available at https://github.com/arturxe2/T-DEED.

📄 PDF Abstract BibTeX arXiv:2404.05392

Code (1)

arturxe2/t-deed 공식 구현 pytorch

Tasks

Decoder

Similar Papers 제목 키워드 기반

Enhancing into the codec: Noise Robust Speech Coding with Vector-Quantized Autoencoders

2021-02-12 · Jonah Casebeer, Vinjai Vale, Umut Isik, Jean-Marc Valin 외

Audio codecs based on discretized neural autoencoders have recently been developed and shown to provide significantly higher compression levels for comparable quality speech output. However, these models are tightly coup…

Towards Robust Video Instance Segmentation with Temporal-Aware Transformer

2023-01-20 · Zhenghao Zhang, Fangtao Shao, Zuozhuo Dai, Siyu Zhu

Most existing transformer based video instance segmentation methods extract per frame features independently, hence it is challenging to solve the appearance deformation problem. In this paper, we observe the temporal in…

DecoderInstance SegmentationSemantic SegmentationVideo Instance Segmentation

DiT-JSCC: Rethinking Deep JSCC with Diffusion Transformers and Semantic Representations

2026-01-06 · Kailin Tan, Jincheng Dai, Sixian Wang, Guo Lu 외 arxiv

Generative joint source-channel coding (GJSCC) has emerged as a new Deep JSCC paradigm for achieving high-fidelity and robust image transmission under extreme wireless channel conditions, such as ultra-low bandwidth and …

3DGS-Enhancer: Enhancing Unbounded 3D Gaussian Splatting with View-consistent 2D Diffusion Priors

2024-10-21 · Xi Liu, Chaoyi Zhou, Siyu Huang

Novel-view synthesis aims to generate novel views of a scene from multiple input images or videos, and recent advancements like 3D Gaussian splatting (3DGS) have achieved notable success in producing photorealistic rende…

3DGSDecoderNovel View SynthesisVideo Generation

S&CNet: Monocular Depth Completion for Autonomous Systems and 3D Reconstruction

2019-07-13 · Lei Zhang, Weihai Chen, Chao Hu, Xingming Wu 외

Dense depth completion is essential for autonomous systems and 3D reconstruction. In this paper, a lightweight yet efficient network (S\&CNet) is proposed to obtain a good trade-off between efficiency and accuracy for th…

3D ReconstructionAutonomous DrivingDecoderDepth Completion