paper-with-me

Papers

Deep Learning Methods for Efficient Large Scale Video Labeling

2017-06-14 · Miha Skalic, Marcin Pekalski, Xingguo E. Pan

We present a solution to "Google Cloud and YouTube-8M Video Understanding Challenge" that ranked 5th place. The proposed model is an ensemble of three model families, two frame level and one video level. The training was performed on augmented dataset, with cross validation.

📄 PDF Abstract BibTeX arXiv:1706.04572

Code (1)

mpekalski/Y8M 공식 구현 tf

Tasks

Deep LearningVideo Understanding

Similar Papers 제목 키워드 기반

Large scale weakly and semi-supervised learning for low-resource video ASR

2020-05-16 · Kritika Singh, Vimal Manohar, Alex Xiao, Sergey Edunov 외

Many semi- and weakly-supervised approaches have been investigated for overcoming the labeling cost of building high quality speech recognition systems. On the challenging task of transcribing social media videos in low-…

Decoderspeech-recognitionSpeech Recognition

A YOLO-Based Semi-Automated Labeling Approach to Improve Fault Detection Efficiency in Railroad Videos

2025-04-01 · Dylan Lester, James Gao, Samuel Sutphin, Pingping Zhu 외

Manual labeling for large-scale image and video datasets is often time-intensive, error-prone, and costly, posing a significant barrier to efficient machine learning workflows in fault detection from railroad videos. Thi…

Fault Detection

VideoPro: A Visual Analytics Approach for Interactive Video Programming

2023-08-01 · Jianben He, Xingbo Wang, Kam Kwai Wong, Xijie Huang 외

Constructing supervised machine learning models for real-world video analysis require substantial labeled data, which is costly to acquire due to scarce domain expertise and laborious manual inspection. While data progra…

Visual Semantic Role Labeling for Video Understanding

2021-04-02 · CVPR 2021 1 · Arka Sadhu, Tanmay Gupta, Mark Yatskar, Ram Nevatia 외

We propose a new framework for understanding and representing related salient events in a video using visual semantic role labeling. We represent videos as a set of related events, wherein each event consists of a verb a…

Semantic Role LabelingVideo RecognitionVideo Understanding

SVD: A Large-Scale Short Video Dataset for Near-Duplicate Video Retrieval

2019-10-01 · ICCV 2019 10 · Qing-Yuan Jiang, Yi He, Gen Li, Jian Lin 외

With the explosive growth of video data in real applications, near-duplicate video retrieval (NDVR) has become indispensable and challenging, especially for short videos. However, all existing NDVR datasets are introduc…

DiversityRetrievalVideo Retrieval