paper-with-me

Papers

BigSmall: Efficient Multi-Task Learning for Disparate Spatial and Temporal Physiological Measurements

2023-03-21 · Girish Narayanswamy, Yujia Liu, Yuzhe Yang, Chengqian Ma, Xin Liu, Daniel McDuff, Shwetak Patel

Understanding of human visual perception has historically inspired the design of computer vision architectures. As an example, perception occurs at different scales both spatially and temporally, suggesting that the extraction of salient visual information may be made more effective by paying attention to specific features at varying scales. Visual changes in the body due to physiological processes also occur at different scales and with modality-specific characteristic properties. Inspired by this, we present BigSmall, an efficient architecture for physiological and behavioral measurement. We present the first joint camera-based facial action, cardiac, and pulmonary measurement model. We propose a multi-branch network with wrapping temporal shift modules that yields both accuracy and efficiency gains. We observe that fusing low-level features leads to suboptimal performance, but that fusing high level features enables efficiency gains with negligible loss in accuracy. Experimental results demonstrate that BigSmall significantly reduces the computational costs. Furthermore, compared to existing task-specific models, BigSmall achieves comparable or better results on multiple physiological measurement tasks simultaneously with a unified model.

📄 PDF Abstract BibTeX arXiv:2303.11573

Code (2)

girishvn/bigsmall 공식 구현 pytorch
ubicomplab/rppg-toolbox pytorch

Tasks

Multi-Task Learning

Similar Papers 제목 키워드 기반

Analysing Fairness of Privacy-Utility Mobility Models

2023-04-10 · Yuting Zhan, Hamed Haddadi, Afra Mashhadi

Preserving the individuals' privacy in sharing spatial-temporal datasets is critical to prevent re-identification attacks based on unique trajectories. Existing privacy techniques tend to propose ideal privacy-utility tr…

FairnessPrivacy PreservingRepresentation Learning

Taming Domain Shift in Multi-source CT-Scan Classification via Input-Space Standardization

2025-07-26 · Chia-Ming Lee, Bo-Cheng Qiu, Ting-Yao Chen, Ming-Han Sun 외 arxiv

Multi-source CT-scan classification suffers from domain shifts that impair cross-source generalization. While preprocessing pipelines combining Spatial-Slice Feature Learning (SSFL++) and Kernel-Density-based Slice Sampl…

A Multi-Level, Multi-Scale Visual Analytics Approach to Assessment of Multifidelity HPC Systems

2023-06-15 · Shilpika, Bethany Lusch, Murali Emani, Filippo Simini 외

The ability to monitor and interpret of hardware system events and behaviors are crucial to improving the robustness and reliability of these systems, especially in a supercomputing facility. The growing complexity and s…

Magnifying Subtle Facial Motions for Effective 4D Expression Recognition

2021-05-05 · Qingkai Zhen, Di Huang, Yunhong Wang, Hassen Drira 외

In this paper, an effective pipeline to automatic 4D Facial Expression Recognition (4D FER) is proposed. It combines two growing but disparate ideas in Computer Vision -- computing the spatial facial deformations using t…

Emotion ClassificationFacial Expression RecognitionFacial Expression Recognition (FER)

Spatial-Temporal-Fusion BNN: Variational Bayesian Feature Layer

2021-12-12 · Shiye Lei, Zhuozhuo Tu, Leszek Rutkowski, Feng Zhou 외

Bayesian neural networks (BNNs) have become a principal approach to alleviate overconfident predictions in deep learning, but they often suffer from scaling issues due to a large number of distribution parameters. In thi…

Adversarial RobustnessUncertainty QuantificationVariational Inference