paper-with-me

홈 › Papers

Understanding Video Content: Efficient Hero Detection and Recognition for the Game "Honor of Kings"

2019-07-18 · Wentao Yao, Zixun Sun, Xiao Chen

In order to understand content and automatically extract labels for videos of the game "Honor of Kings", it is necessary to detect and recognize characters (called "hero") together with their camps in the game video. In this paper, we propose an efficient two-stage algorithm to detect and recognize heros in game videos. First, we detect all heros in a video frame based on blood bar template-matching method, and classify them according to their camps (self/ friend/ enemy). Then we recognize the name of each hero using one or more deep convolution neural networks. Our method needs almost no work for labelling training and testing samples in the recognition stage. Experiments show its efficiency and accuracy in the task of hero detection and recognition in game videos.

📄 PDF Abstract BibTeX arXiv:1907.07854

Code (0)

등록된 구현이 없습니다.

Tasks

Template Matching

Methods 이 논문이 사용한 방법론

Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…

Similar Papers 제목 키워드 기반

HERO: Hierarchical Encoder for Video+Language Omni-representation Pre-training

2020-05-01 · EMNLP 2020 11 · Linjie Li, Yen-Chun Chen, Yu Cheng, Zhe Gan 외

We present HERO, a novel framework for large-scale video+language omni-representation learning. HERO encodes multimodal inputs in a hierarchical structure, where local context of a video frame is captured by a Cross-moda…

Language ModelingLanguage ModellingMasked Language ModelingMoment Retrieval+7

Detection and Analysis of Content Creator Collaborations in YouTube Videos using Face- and Speaker-Recognition

2018-07-05 · Moritz Lode, Michael Örtl, Christian Koch, Amr Rizk 외

This work discusses and implements the application of speaker recognition for the detection of collaborations in YouTube videos. CATANA, an existing framework for detection and analysis of YouTube collaborations, is util…

Active Speaker DetectionFace RecognitionSpeaker Recognition

Are you a hero or a villain? A semantic role labelling approach for detecting harmful memes.

2022-05-01 · CONSTRAINT (ACL) 2022 5 · Shaik Fharook, Syed Sufyan Ahmed, Gurram Rithika, Sumith Sai Budde 외

Identifying good and evil through representations of victimhood, heroism, and villainy (i.e., role labeling of entities) has recently caught the research community’s interest. Because of the growing popularity of memes, …

named-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)NER+1

Do Current Video LLMs Have Strong OCR Abilities? A Preliminary Study

2024-12-29 · Yulin Fei, Yuhui Gao, Xingyuan Xian, Xiaojin Zhang 외

With the rise of multimodal large language models, accurately extracting and understanding textual information from video content, referred to as video based optical character recognition (Video OCR), has become a crucia…

Motion DetectionOptical Character RecognitionOptical Character Recognition (OCR)Temporal Localization

RadarLCD: Learnable Radar-based Loop Closure Detection Pipeline

2023-09-13 · Mirko Usuelli, Matteo Frosi, Paolo Cudrano, Simone Mentasti 외

Loop Closure Detection (LCD) is an essential task in robotics and computer vision, serving as a fundamental component for various applications across diverse domains. These applications encompass object recognition, imag…

Image RetrievalLoop Closure DetectionObject RecognitionRadar odometry