paper-with-me

홈 › Papers

Local-Global Associative Frame Assemble in Video Re-ID

2021-10-22 · Qilei Li, Jiabo Huang, Shaogang Gong

Noisy and unrepresentative frames in automatically generated object bounding boxes from video sequences cause significant challenges in learning discriminative representations in video re-identification (Re-ID). Most existing methods tackle this problem by assessing the importance of video frames according to either their local part alignments or global appearance correlations separately. However, given the diverse and unknown sources of noise which usually co-exist in captured video data, existing methods have not been effective satisfactorily. In this work, we explore jointly both local alignments and global correlations with further consideration of their mutual promotion/reinforcement so to better assemble complementary discriminative Re-ID information within all the relevant frames in video tracklets. Specifically, we concurrently optimise a local aligned quality (LAQ) module that distinguishes the quality of each frame based on local alignments, and a global correlated quality (GCQ) module that estimates global appearance correlations. With the help of a local-assembled global appearance prototype, we associate LAQ and GCQ to exploit their mutual complement. Extensive experiments demonstrate the superiority of the proposed model against state-of-the-art methods on five Re-ID benchmarks, including MARS, Duke-Video, Duke-SI, iLIDS-VID, and PRID2011.

📄 PDF Abstract BibTeX arXiv:2110.12018

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

X2Video: Adapting Diffusion Models for Multimodal Controllable Neural Video Rendering

2025-10-09 · Zhitong Huang, Mohan Zhang, Renhan Wang, Rui Tang 외 arxiv

We present X2Video, the first diffusion model for rendering photorealistic videos guided by intrinsic channels including albedo, normal, roughness, metallicity, and irradiance, while supporting intuitive multi-modal cont…

Video GenerationImage Generation

Hierarchical Associative Memory

2021-07-14 · Dmitry Krotov

Dense Associative Memories or Modern Hopfield Networks have many appealing properties of associative memory. They can do pattern completion, store a large number of memories, and can be described using a recurrent neural…

Deep Image Spatial Transformation for Person Image Generation

2020-03-02 · CVPR 2020 6 · Yurui Ren, Xiaoming Yu, Junming Chen, Thomas H. Li 외

Pose-guided person image generation is to transform a source person image to a target pose. This task requires spatial manipulations of source data. However, Convolutional Neural Networks are limited by the lack of abili…

Image Generation

Deep Spatial Transformation for Pose-Guided Person Image Generation and Animation

2020-08-27 · Yurui Ren, Ge Li, Shan Liu, Thomas H. Li

Pose-guided person image generation and animation aim to transform a source person image to target poses. These tasks require spatial manipulation of source data. However, Convolutional Neural Networks are limited by the…

Image AnimationImage GenerationNovel View Synthesis

Firing Rate Models as Associative Memory: Excitatory-Inhibitory Balance for Robust Retrieval

2024-11-11 · Simone Betteti, Giacomo Baggio, Francesco Bullo, Sandro Zampieri

Firing rate models are dynamical systems widely used in applied and theoretical neuroscience to describe local cortical dynamics in neuronal populations. By providing a macroscopic perspective of neuronal activity, these…

Retrieval