paper-with-me

홈 › Papers

CanonSLR: Canonical-View Guided Multi-View Continuous Sign Language Recognition

2026-04-20 · Xu Wang, Shengeng Tang, Wan Jiang, Yaxiong Wang, Lechao Cheng, Richang Hong arxiv

Continuous Sign Language Recognition (CSLR) has achieved remarkable progress in recent years; however, most existing methods are developed under single-view settings and thus remain insufficiently robust to viewpoint variations in real-world scenarios. To address this limitation, we propose CanonSLR, a canonical-view guided framework for multi-view CSLR. Specifically, we introduce a frontal-view-anchored teacher-student learning strategy, in which a teacher network trained on frontal-view data provides canonical temporal supervision for a student network trained on all viewpoints. To further reduce cross-view semantic discrepancy, we propose Sequence-Level Soft-Target Distillation, which transfers structured temporal knowledge from the frontal view to non-frontal samples, thereby alleviating gloss boundary ambiguity and category confusion caused by occlusion and projection variation. In addition, we introduce Temporal Motion Relational Enhancement to explicitly model motion-aware temporal relations in high-level visual features, strengthening stable dynamic representations while suppressing viewpoint-sensitive appearance disturbances. To support multi-view CSLR research, we further develop a universal multi-view sign language data construction pipeline that transforms original single-view RGB videos into semantically consistent, temporally coherent, and viewpoint-controllable multi-view sign language videos. Based on this pipeline, we extend PHOENIX-2014T and CSL-Daily into two seven-view benchmarks, namely PT14-MV and CSL-MV, providing a new experimental foundation for multi-view CSLR. Extensive experiments on PT14-MV and CSL-MV demonstrate that CanonSLR consistently outperforms existing approaches under multi-view settings and exhibits stronger robustness, especially on challenging non-frontal views.

📄 PDF Abstract BibTeX arXiv:2604.18184

Code (0)

등록된 구현이 없습니다.

Tasks

Sign Language Recognition

Similar Papers 제목 키워드 기반

Learning Stable Canonical Worlds for Novel View Synthesis and Beyond

2026-06-22 · Xiaoyu Xu, Jian Zou, Sheyang Tang, Zhihua Wang 외 arxiv

Feed-forward Gaussian splatting (FFGS) facilitates real-time novel view synthesis, yet current methods often remain tied to view-dependent predictions. As more input views are added, they may accumulate noisy or redundan…

Semantic SegmentationNovel View Synthesis

Learning Canonical View Representation for 3D Shape Recognition with Arbitrary Views

2021-08-16 · ICCV 2021 10 · Xin Wei, Yifei Gong, Fudong Wang, Xing Sun 외

In this paper, we focus on recognizing 3D shapes from arbitrary views, i.e., arbitrary numbers and positions of viewpoints. It is a challenging and realistic setting for view-based 3D shape recognition. We propose a cano…

3D Shape Recognition3D Shape Representation

Constructing Canonical Regions for Fast and Effective View Selection

2016-06-01 · CVPR 2016 6 · Wencheng Wang, Tianhao Gao

In view selection, little work has been done for optimizing the search process; views must be densely distributed and checked individually. Thus, evaluating poor views wastes much time, and a poor view may even be miside…

Canonical Correlation Guided Deep Neural Network

2024-09-28 · Zhiwen Chen, Siwen Mo, Haobin Ke, Steven X. Ding 외

Learning representations of two views of data such that the resulting representations are highly linearly correlated is appealing in machine learning. In this paper, we present a canonical correlation guided learning fra…

Fault DiagnosisRepresentation Learning

Learning about Canonical Views from Internet Image Collections

2012-12-01 · NeurIPS 2012 12 · Elad Mezuman, Yair Weiss

Although human object recognition is supposedly robust to viewpoint, much research on human perception indicates that there is a preferred or “canonical” view of objects. This phenomenon was discovered more than 30 years…

Object Recognition