Attention-based Dual-stream Vision Transformer for Radar Gait Recognition
Radar gait recognition is robust to light variations and less infringement on privacy. Previous studies often utilize either spectrograms or cadence velocity diagrams. While the former shows the time-frequency patterns, the latter encodes the repetitive frequency patterns. In this work, a dual-stream neural network with attention-based fusion is proposed to fully aggregate the discriminant information from these two representations. The both streams are designed based on the Vision Transformer, which well captures the gait characteristics embedded in these representations. The proposed method is validated on a large benchmark dataset for radar gait recognition, which shows that it significantly outperforms state-of-the-art solutions.
Code (0)
등록된 구현이 없습니다.
Tasks
Gait RecognitionMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
physfusion: A Transformer-based Dual-Stream Radar and Vision Fusion Framework for Open Water Surface Object Detection
Detecting water-surface targets for Unmanned Surface Vehicles (USVs) is challenging due to wave clutter, specular reflections, and weak appearance cues in long-range observations. Although 4D millimeter-wave radar comple…
Object DetectionPoint CloudsRCBEVDet: Radar-camera Fusion in Bird's Eye View for 3D Object Detection
Three-dimensional object detection is one of the key tasks in autonomous driving. To reduce costs in practice, low-cost multi-view cameras for 3D object detection are proposed to replace the expansive LiDAR sensors. Howe…
3D Object Detection3D Object Detection (RoI)Autonomous DrivingObject+3TransCAR: Transformer-based Camera-And-Radar Fusion for 3D Object Detection
Despite radar's popularity in the automotive industry, for fusion-based 3D object detection, most existing works focus on LiDAR and camera fusion. In this paper, we propose TransCAR, a Transformer-based Camera-And-Radar …
3D Object DetectionDecoderObjectobject-detection+1CORENet: Cross-Modal 4D Radar Denoising Network with LiDAR Supervision for Autonomous Driving
4D radar-based object detection has garnered great attention for its robustness in adverse weather conditions and capacity to deliver rich spatial information across diverse driving scenarios. Nevertheless, the sparse an…
Autonomous DrivingObject DetectionPoint CloudsDual-Stream Attention Transformers for Sewer Defect Classification
We propose a dual-stream multi-scale vision transformer (DS-MSHViT) architecture that processes RGB and optical flow inputs for efficient sewer defect classification. Unlike existing methods that combine the predictions …
ClassificationOptical Flow Estimation