paper-with-me

홈 › Papers

CAST: Channel-Aware Spatial Transfer Learning with Pseudo-Image Radar for Sign Language Recognition

2026-05-09 · Md. Shakhoyat Rahman Shujon, Sheikh Md. Galib Mahim, Md. Milon Islam, Md Rezwanul Haque, Md Rabiul Islam, Hamdi Altaheri, Fakhri Karray arxiv

We propose CAST, a dual-stream architecture that utilizes channel-aware spatial transfer learning for isolated sign language recognition addressing the challenges of magnitude-only 60~GHz radar Range-Time Maps (RTM). The proposed framework combines three physics-aware architectures with pretrained vision backbones, which operate under radar-only constraints across clinical and alphabetical gestures. First, an explicit decibel-to-linear inversion is combined with a windowed fast Fourier transform that extracts Cadence Velocity Diagrams (CVD) while avoiding the harmonic artifacts that arise from the spectral analysis of log-compressed signals. Second, a cross-antenna spatial attention module applies attention to raw antenna channels before the convolution, preserving inter-receiver amplitude covariance. Third, an asymmetric cross-attention mechanism fuses representations from parallel ConvNeXt-Tiny (CVD) and EfficientNetV2-S (RTM) backbones. Extensive experiments reveal that the architecture achieves a Top-1 accuracy of 80.5% under 5-fold cross-validation, establishing a 3.3% improvement over the best single-model baseline (77.2%). The findings suggest that physics-aware signal representations form a promising direction for radar-only sign language recognition under constrained sensor modalities. The source code is available at: https://github.com/Shakhoyat/CAST-at-SignEval2026.

📄 PDF Abstract BibTeX arXiv:2605.08663

Code (0)

등록된 구현이 없습니다.

Tasks

Sign Language RecognitionTransfer Learning

Similar Papers 제목 키워드 기반

Self-training Room Layout Estimation via Geometry-aware Ray-casting

2024-07-21 · Bolivar Solarte, Chin-Hsuan Wu, Jin-Cheng Jhang, Jonathan Lee 외

In this paper, we introduce a novel geometry-aware self-training framework for room layout estimation models on unseen scenes with unlabeled data. Our approach utilizes a ray-casting formulation to aggregate multiple est…

Room Layout Estimation

EVM Mitigation with PAPR and ACLR Constraints in Large-Scale MIMO-OFDM Using TOP-ADMM

2022-05-25 · Shashi Kant, Mats Bengtsson, Gabor Fodor, Bo Göransson 외

Although signal distortion-based peak-to-average power ratio (PAPR) reduction is a feasible candidate for orthogonal frequency division multiplexing (OFDM) to meet standard/regulatory requirements, the error vector magni…

Pseudo-Zernike Based Multi-Pass Automatic Target Recognition From Multi-Channel SAR

2014-04-07 · Carmine Clemente, Luca Pallotta, Ian Proudler, Antonio De Maio 외

The capability to exploit multiple sources of information is of fundamental importance in a battlefield scenario. Information obtained from different sources, and separated in space and time, provide the opportunity to e…

Diversity

Tropospheric temperature and humidity profile retrieval from Meteosat Flexible Combined Imager based on deep learning

2026-08-26 · Alejandro Salgueiro, Johannes Rausch, Julie Thérèse Villinger, Angela Meyer arxiv

The Meteosat Third Generation (MTG) Flexible Combined Imager (FCI) offers new opportunities for tropospheric temperature and humidity profiling, at higher spatio-temporal resolutions and expanded spectral coverage relati…

User Subgrouping and Power Control for Multicast Massive MIMO over Spatially Correlated Channels

2024-09-18 · Alejandro de la Fuente, Giovanni Interdonato, Giuseppe Araniti

Massive multiple-input-multiple-output (MIMO) is unquestionably a key enabler of the fifth-generation (5G) technology for mobile systems, enabling to meet the high requirements of upcoming mobile broadband services. Phys…

Fairness