paper-with-me

홈 › Papers

Rotation-Constrained Cross-View Feature Fusion for Multi-View Appearance-based Gaze Estimation

2023-05-22 · Yoichiro Hisadome, Tianyi Wu, Jiawei Qin, Yusuke Sugano

Appearance-based gaze estimation has been actively studied in recent years. However, its generalization performance for unseen head poses is still a significant limitation for existing methods. This work proposes a generalizable multi-view gaze estimation task and a cross-view feature fusion method to address this issue. In addition to paired images, our method takes the relative rotation matrix between two cameras as additional input. The proposed network learns to extract rotatable feature representation by using relative rotation as a constraint and adaptively fuses the rotatable features via stacked fusion modules. This simple yet efficient approach significantly improves generalization performance under unseen head poses without significantly increasing computational cost. The model can be trained with random combinations of cameras without fixing the positioning and can generalize to unseen camera pairs during inference. Through experiments using multiple datasets, we demonstrate the advantage of the proposed method over baseline methods, including state-of-the-art domain generalization approaches. The code will be available at https://github.com/ut-vision/Rot-MVGaze.

📄 PDF Abstract BibTeX arXiv:2305.12704

Code (1)

ut-vision/rot-mvgaze 공식 구현 pytorch

Tasks

Domain GeneralizationGaze Estimation

Similar Papers 제목 키워드 기반

Deep Fusion of Ultra-Low-Resolution Thermal Camera and Gyroscope Data for Lighting-Robust and Compute-Efficient Rotational Odometry

2025-06-14 · Farida Mohsen, Ali Safa

Accurate rotational odometry is crucial for autonomous robotic systems, particularly for small, power-constrained platforms such as drones and mobile robots. This study introduces thermal-gyro fusion, a novel sensor fusi…

Computational EfficiencySensor Fusion

DRKF: Distilled Rotated Kernel Fusion for Efficient Rotation Invariant Descriptors in Local Feature Matching

2022-09-22 · Ranran Huang, Jiancheng Cai, Chao Li, Zhuoyuan Wu 외

The performance of local feature descriptors degrades in the presence of large rotation variations. To address this issue, we present an efficient approach to learning rotation invariant descriptors. Specifically, we pro…

Knowledge Distillation

Multi-View Industrial Anomaly Detection with Epipolar Constrained Cross-View Fusion

2025-03-14 · Yifan Liu, Xun Xu, Shijie Li, Jingyi Liao 외

Multi-camera systems provide richer contextual information for industrial anomaly detection. However, traditional methods process each view independently, disregarding the complementary information across viewpoints. Exi…

Anomaly Detection

Wide-Area Crowd Counting: Multi-View Fusion Networks for Counting in Large Scenes

2020-12-02 · Qi Zhang, Antoni B. Chan

Crowd counting in single-view images has achieved outstanding performance on existing counting datasets. However, single-view counting is not applicable to large and wide scenes (e.g., public parks, long subway platforms…

Crowd Counting

A Cross-view Fusion Framework for Robust 6-DoF Grasp Pose Estimation

2026-06-05 · Kangjian Zhu, Haobo Jiang, Jianjun Qian, Jin Xie arxiv

In this paper, we propose a cross-view fusion framework that enhances the robustness of 6-DoF grasp pose estimation in corner views. Our framework alleviates occlusion by incorporating an auxiliary view and avoids the ti…

Contrastive LearningPose Estimation