Understanding Video Content: Efficient Hero Detection and Recognition for the Game "Honor of Kings"
In order to understand content and automatically extract labels for videos of the game "Honor of Kings", it is necessary to detect and recognize characters (called "hero") together with their camps in the game video. In this paper, we propose an efficient two-stage algorithm to detect and recognize heros in game videos. First, we detect all heros in a video frame based on blood bar template-matching method, and classify them according to their camps (self/ friend/ enemy). Then we recognize the name of each hero using one or more deep convolution neural networks. Our method needs almost no work for labelling training and testing samples in the recognition stage. Experiments show its efficiency and accuracy in the task of hero detection and recognition in game videos.
Code (0)
등록된 구현이 없습니다.
Tasks
Template MatchingMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
HERO: Hierarchical Encoder for Video+Language Omni-representation Pre-training
We present HERO, a novel framework for large-scale video+language omni-representation learning. HERO encodes multimodal inputs in a hierarchical structure, where local context of a video frame is captured by a Cross-moda…
Language ModelingLanguage ModellingMasked Language ModelingMoment Retrieval+7Detection and Analysis of Content Creator Collaborations in YouTube Videos using Face- and Speaker-Recognition
This work discusses and implements the application of speaker recognition for the detection of collaborations in YouTube videos. CATANA, an existing framework for detection and analysis of YouTube collaborations, is util…
Active Speaker DetectionFace RecognitionSpeaker RecognitionAre you a hero or a villain? A semantic role labelling approach for detecting harmful memes.
Identifying good and evil through representations of victimhood, heroism, and villainy (i.e., role labeling of entities) has recently caught the research community’s interest. Because of the growing popularity of memes, …
named-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)NER+1Do Current Video LLMs Have Strong OCR Abilities? A Preliminary Study
With the rise of multimodal large language models, accurately extracting and understanding textual information from video content, referred to as video based optical character recognition (Video OCR), has become a crucia…
Motion DetectionOptical Character RecognitionOptical Character Recognition (OCR)Temporal LocalizationRadarLCD: Learnable Radar-based Loop Closure Detection Pipeline
Loop Closure Detection (LCD) is an essential task in robotics and computer vision, serving as a fundamental component for various applications across diverse domains. These applications encompass object recognition, imag…
Image RetrievalLoop Closure DetectionObject RecognitionRadar odometry